InternRobotics / InternRobotics/MeshCoder

Release MeshCoder artifacts (models, dataset) on Hugging Face

Aperta
#1 0 commenti 1 reazione 0 assegnatari Vedi su GitHub

Nessuno ha ancora preso questa issue.

Lingua principale
Jupyter Notebook
Stelle
534
Fork
29
Metriche di merge delle PR
Nessuna PR unita negli ultimi 30g

Descrizione

Hi @ZhaoyangLyu 🤗

Niels here from the open-source team at Hugging Face. I discovered your work through Hugging Face's daily papers as yours got featured: https://huggingface.co/papers/2508.14879.
The paper page lets people discuss about your paper and lets them find artifacts about it (your models, datasets or demo for instance), you can also claim
the paper as yours which will show up on your public profile at HF, add Github and project page URLs.

I saw on your GitHub repository that you plan to release the code for MeshCoder and the associated large-scale paired object-code dataset by November 2025. That's fantastic news!
It'd be great to make these checkpoints and the dataset available on the 🤗 hub once they are released, to improve their discoverability/visibility.
We can add tags so that people find them when filtering https://huggingface.co/models and https://huggingface.co/datasets. The MeshCoder model (a multimodal LLM that translates 3D point clouds into Blender Python scripts) would likely fit an "image-to-3d" pipeline tag, and the "large-scale paired object-code dataset" would have an "image-to-3d" task category.

Uploading models

See here for a guide: https://huggingface.co/docs/hub/models-uploading.

In this case, we could leverage the PyTorchModelHubMixin class which adds from_pretrained and push_to_hub to any custom nn.Module. Alternatively, one can leverages the hf_hub_download one-liner to download a checkpoint from the hub.

We encourage researchers to push each model checkpoint to a separate model repository, so that things like download stats also work. We can then also link the checkpoints to the paper page.

Uploading dataset

Would be awesome to make the dataset available on 🤗 , so that people can do:

from datasets import load_dataset

dataset = load_dataset("your-hf-org-or-username/your-dataset")

See here for a guide: https://huggingface.co/docs/datasets/loading.

Besides that, there's the dataset viewer which allows people to quickly explore the first few rows of the data in the browser.

Let me know if you're interested/need any help regarding this once the artifacts are ready for release!

Cheers,

Niels
ML Engineer @ HF 🤗

Guida per i contributori

Nessuna guida per i contributori indicizzata per questo repository

Come iniziare

  1. Leggi tutta la issue e poi la guida ai contributi del progetto.
  2. Commenta sulla issue per dire che te ne occupi tu — evita che due persone facciano lo stesso lavoro.
  3. Fai un fork del repository e lavora su un branch.
  4. Apri una pull request che faccia riferimento al numero della issue.

Direzione di ricerca

Inizia verificando quali checkpoint di MeshCoder e quale dataset di object-code associato sono pronti per il rilascio. Consulta le guide di Hugging Face per il caricamento dei modelli e il caricamento dei dataset, inclusi PyTorchModelHubMixin e load_dataset. Il lavoro è completato quando ogni checkpoint ha un repository di modello separato, il dataset è disponibile sull’Hub per essere caricato e visualizzato e gli artefatti possono essere collegati all’articolo.

Scritto dal modello di indicizzazione a partire dal testo della issue.

Valutazione

Stack tecnologico
blender, huggingface, python, pytorch
Ambito
data, machine-learning, release
Tipo di issue
Funzionalità
Difficoltà
4/5
Tempo stimato
3-5 giorni
Stato di attività
Ferma
Chiarezza
Abbastanza chiara
Idoneità per principianti
35/100

Ricevi le nuove issue nella tua casella

Un breve riepilogo di issue GitHub adatte ai principianti.