NVIDIA / NVIDIA/cuvs

[FEA] KMeans - Within Iteration Inertia Computation Should Account for Sample Weights

Open
#1,940 0 comments 0 reactions 0 assignees View on GitHub

Nobody has claimed this yet.

feature request
Dominant language
Cuda
Stars
854
Forks
236
Avg merge
3d 3h
Merged PRs (30d)
62

Description

Final inertia computation is weighted for both kmeans and batched kmeans (#1886), but the within iteration computation (to check for convergence with inertia_check) seems to be not weighted.

Contributor guide

Open the contributing guide

First steps

  1. Read the whole issue, then the project's contributing guide.
  2. Comment on the issue to say you are picking it up — it saves two people doing the same work.
  3. Fork the repository and make your change on a branch.
  4. Open a pull request that references the issue number.

Research direction

No files or tests are named. Start by tracing the KMeans and batched KMeans convergence path used by inertia_check, then compare its within-iteration inertia calculation with the weighted final inertia added in #1886. Done means convergence checks account for sample weights and coverage exists for weighted KMeans cases.

Written by the indexing model from the issue text.

Assessment

Domain
machine-learning
Issue type
Feature
Difficulty
4/5
Estimated time
3-5 days
Activity status
Quiet
Clarity
Mostly clear
Newbie friendliness
48/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.