ICML 2025PastOther
Championing Open-source DEvelopment in ML Workshop @ ICML25
CODEML@ICML25
- Submission deadline
- May 27, 2025, 11:59 UTCimported from OpenReview — check the website for extensions
- Submission portal
- OpenReview
- Notes
- Topics were auto-suggested and may be imprecise — edits welcome.
Accepted papers (44)
Fetched from OpenReview (v2) on 2026-06-10.
$\texttt{markovml}$: A Python Package for Verifying Markov Processes with Embedded Machine Learning Models
A2Perf: Benchmarking Autonomous Agents End-to-End in Realistic Domains
AIF-GEN: Open-Source Platform and Synthetic Dataset Suite for Reinforcement Learning on Large Language Models
An LLM-Powered Tool for Enhancing Scientific Open-Source Repositories
An Open-Source Software Toolkit & Benchmark Suite for the Evaluation and Adaptation of Multimodal Action Models
Bencher: Simple and Reproducible Benchmarking for Black-Box Optimization
BoFire: Bayesian Optimization Framework Intended for Real Experiments
Common Task Framework For a Critical Evaluation of Scientific Machine Learning Algorithms
Control Flow Operators in PyTorch
cp_measure: API-first feature extraction for image-based profiling workflows
DeepChem-Variant: A Modular Open Source Framework for Genomic Variant Calling
Deploying User-Friendly Software: Six Recommendations to Make Single-Cell Foundation Models More Usable For Scientific Discovery
Developing and Maintaining an Open-Source Repository of AI Evaluations: Challenges and Insights
DINOHash: Learning Adversarially Robust Perceptual Hashes from Self-Supervised Features
DISCO: A Browser-Based Privacy-Preserving Framework for Distributed Collaborative Learning
EXO Gym: a simulation environment for low-bandwidth training
FedRAG: A Framework for Fine-Tuning Retrieval-Augmented Generation Systems
If open source is to win, it must go public
KernelBot: A Competition Platform for Writing Heterogeneous GPU Code
laplax - Laplace Approximations with JAX
Liger-Kernel: Efficient Triton Kernels for LLM Training
LUQ: Language Models Uncertainty Quantification Toolkit
M(M)ORE : Massive Multimodal Open RAG & Extraction
Maintaining MTEB: Towards Long Term Usability and Reproducibility of Embedding Benchmarks
Meta-World+: An Improved, Standardized, RL Benchmark
N$^2$: A Unified Python Package and Test Bench for Nearest Neighbor-Based Matrix Completion
olmOCR: Unlocking Trillions of Tokens in PDFs with Vision Language Models
Open-Source Foosball Benchmark for Deep Reinforcement Learning
Orthogonium: A Unified, Efficient Library of Orthogonal and 1‑Lipschitz Building Blocks
Provenance Design and Evolution in a Production ML Library
PyLO: Towards Accessible Learned Optimizers in Pytorch
RepoST: Scalable Repository-Level Coding Environment Construction with Sandbox Testing
Reproducible sampling from intractable distributions with Pigeons.jl
SAGDA: Open-Source Synthetic Agriculture Data for Africa
Scaling Private Deep Learning with Opacus: Advances for Large Language Models
skglm: Improving scikit-learn for Regularized Generalized Linear Models
Spatial Reasoners for Continuous Variables in Any Domain
Swizz: One-Liner Figures, LaTeX Tables, and Flexible Layouts for Scientific Papers
TGM: A Modular Framework for Machine Learning on Temporal Graphs
TorchAO: PyTorch-Native Training-to-Serving Model Optimization
TorchTitan: A PyTorch Native Platform for Training Generative AI Models
Vulnerability of Text-Matching in ML/AI Conference Reviewer Assignments to Collusions
Write Code that People Want to Use
ZKLoRA: Efficient Zero-Knowledge Proofs for LoRA Verification