Skip to content

Fix ID processing (int->str) - #12

Merged
sfluegel05 merged 2 commits into
devfrom
fix/id-processing
Sep 23, 2026
Merged

sfluegel05 merged 2 commits into
devfrom
fix/id-processing

Conversation

@sfluegel05

Copy link
Copy Markdown
Contributor

IDs are handled as strings in chebILP. However, pandas handles them as numbers by default unless we tell pandas not to.

This issue was also addressed in 1ba9d51, but only for the ILPProblemBuilder, not generally (which did not matter so far since it is the only place using the splits). This PR implements the cleaner solution and addresses the problem right at the parsing stage.

Also, we now require python >=3.11 since networkx does not work with 3.10.

Copilot AI left a comment

Copy link
Copy Markdown

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Copilot review overview

🟢 Approval recommended

No unresolved blocking issues were identified.

Review effort: Lite
Findings: None

What changed in this PR

This PR preserves ChEBI IDs as strings during parsing, corrects split-based filtering, and raises the minimum Python version to 3.11.

Changes:

  • Parse split IDs as strings.
  • Filter samples using split IDs.
  • Require Python 3.11+.
File Description
pyproject.toml Raises the minimum Python version.
chebILP/​molecule_processing/​data_preparation.py Preserves split IDs as strings.
chebILP/​ilp_problem_builder.py Corrects split-based sample filtering.

💡 Add a code-review agent skill or configure MCP servers for context-aware, tailored reviews. Learn more in the docs.

@sfluegel05
sfluegel05 merged commit c9c87ca into dev Sep 23, 2026
1 check passed
@sfluegel05
sfluegel05 deleted the fix/id-processing branch September 23, 2026 11:23
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

2 participants