CLAIRE Lab @EPFL
Pinned Loading
Repositories
- quantile-reward-policy-optimization Public
Official codebase for "Quantile Reward Policy Optimization: Alignment with Pointwise Regression and Exact Partition Functions" (Matrenok et al. 2025).
- python-ml-research-template Public template
A template for starting reproducible Python machine-learning projects with hardware acceleration. Find an example at https://github.com/CLAIRE-Labo/no-representation-no-trust
-
- no-representation-no-trust Public
Codebase to fully reproduce the results of "No Representation, No Trust: Connecting Representation, Collapse, and Trust Issues in PPO" (Moalla et al. 2024). Uses TorchRL and provides extensive tools for studying representation dynamics in policy optimization.
-
- tunable-morl-public Public
Supplementary code for "In Search for Architectures and Loss Functions in Multi-Objective Reinforcement Learning"
- StructuredFFN Public
The official code of "Building on Efficient Foundations: Effectively Training LLMs with Structured Feedforward Layers"
Top languages
Loading…
Most used topics
Loading…