alibaba / alibaba/DataX

从oracle读取数据写入TXT文件,有时候写入的数据变成utf-16格式乱码了,配置的是UTF8格式

Open
#1,240 0 comments 0 reactions 0 assignees View on GitHub
Dominant language
Java
Stars
17.4k
Forks
5.7k
PR merge metrics
No merged PRs in 30d

Description

数据没有特殊字符,乱码的源数据加一个字符就会正常,再加一个字符又会乱码,无限循环

Contributor guide

No contributing guide indexed for this repository

Research direction

The report names no files, tests, or entry points. Start by reproducing the Oracle-to-TXT export with UTF-8 using inputs of different lengths, then inspect the encoding of the resulting files; done means the output remains valid UTF-8 regardless of data length.

Written by the indexing model from the issue text.

Assessment

Tech stack
java
Domain
data-engineering
Issue type
Bug
Difficulty
4/5
Estimated time
3-5 days
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
25/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.