objectionary / objectionary/lints

`incorrect-bytes-format` XSL transformation is too slow on large XMIR files

Open
#984 0 comments 0 reactions 1 assignee View on GitHub

@volodya-lombrozo is already working on this.

Since Jun 24, 2026.

bug
Dominant language
Java
Stars
14
Forks
39
Avg merge
22h 54m
Merged PRs (30d)
90

Description

I'm getting performance warnings when running lints against eo-runtime:

[WARNING] XSL transformation 'incorrect-bytes-format' took 174ms, whereas threshold is 100ms

What happens: The selector //o[normalize-space(string-join(text(), '')) != ''] materializes and concatenates all text node content for every <o> element in the document. Then for each match, matches($bytes, '^(--|[0-9A-F]{2}(-|(-[0-9A-F]{2})+))$') applies a complex alternation regex. On large XMIR with many byte-carrying elements the string-join + normalize-space + regex evaluation adds up significantly.

What should happen: The transformation completes within the 100ms threshold. Using //o[text()] as the initial selector (elements with any text node) before materializing would narrow the candidate set. The regex itself could also be simplified.

Contributor guide

No contributing guide indexed for this repository

First steps

  1. Read the whole issue, then the project's contributing guide.
  2. Comment on the issue to say you are picking it up — it saves two people doing the same work.
  3. Fork the repository and make your change on a branch.
  4. Open a pull request that references the issue number.

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.