queryverse / queryverse/Query.jl
Feature Request - Normalize Names
Nobody has claimed this yet.
- Dominant language
- Julia
- Stars
- 403
- Forks
- 48
- Avg merge
- 3d 6h
- Merged PRs (30d)
- 6
Description
CSV.jl has a normalizenames argument that can be set to true when reading a file. This option replaces invalid identifier characters (spaces) with underscores. I think it would be nice to add this sort of a feature to Query, but I would take it a step further and remove all trailing/leading whitespaces (rather than replacing them with underscores). From the CSV.jl documentation:
"When a column name is not a single atom Julia identifier, this is inconvenient, because f.column one is not valid, so I would have to manually call getproperty(f, Symbol("column one")"
Julia's built-in strip and replace functions should do the job. I'd love to make an attempt to write this myself if you can provide a basic roadmap for me to get started.
Thanks!!
Contributor guide
First steps
- Read the whole issue, then the project's contributing guide.
- Comment on the issue to say you are picking it up — it saves two people doing the same work.
- Fork the repository and make your change on a branch.
- Open a pull request that references the issue number.
Research direction
Start by locating where Query handles or exposes column names, then compare the requested behavior with CSV.jl's normalizenames documentation. Investigate Julia's strip and replace functions for removing leading and trailing whitespace and normalizing invalid identifier characters. Done means normalized column names can be used without manually calling getproperty with a Symbol.
Written by the indexing model from the issue text.
Assessment
- Tech stack
- julia
- Domain
- data
- Issue type
- Feature
- Difficulty
- 3/5
- Estimated time
- 1-2 days
- Activity status
- Stale
- Clarity
- Mostly clear
- Newbie friendliness
- 35/100