opd
Here are 15 public repositories matching this topic...
A curated collection of papers and resources on On-Policy Distillation for Large Language Models.
-
Updated
Aug 12, 2026 - Python
🔥 DanceOPD: On-Policy Generative Field Distillation
-
Updated
Aug 15, 2026 - Python
Source code of paper "RLCSD: Reinforcement Learning with Contrastive On-Policy Self-Distillation"
-
Updated
Aug 21, 2026 - Python
Official implementation of On-Policy Delta Distillation (OPD2)
-
Updated
Aug 7, 2026 - Python
Tiny-R2: A hybrid architecture integrating SWA, CSA, HCA, mHC, and DSMoE under the DeepSeek V4 design paradigm, enabling single-GPU OPD post-training.
-
Updated
May 30, 2026 - Python
The implementation for our paper: TurnSight: Turn-Level Hindsight Self-Distillation for Tool-Integrated Reasoning.
-
Updated
Aug 5, 2026 - Python
H-OPD: Confidence Aware Heterogeneous Multi-Teacher Multimodal On-policy Distillation
-
Updated
Jul 8, 2026 - Python
🔥 Learning When to Trust via Selective Context Preference Optimization
-
Updated
Aug 12, 2026 - Python
Open Source Health Information System. Get your Frappe Health instance ready within a few seconds and uncover the potentials of digitising your operations.
-
Updated
Feb 12, 2024 - Python
Open Source Health Information System. Get your Frappe Health instance ready within a few seconds and uncover the potentials of digitising your operations.
-
Updated
Feb 12, 2024 - Python
This project was realised during our Industrial Project. Real project : https://github.com/Victor804/keithley-software
-
Updated
Aug 29, 2023 - Python
Improve this page
Add a description, image, and links to the opd topic page so that developers can more easily learn about it.
Add this topic to your repo
To associate your repository with the opd topic, visit your repo's landing page and select "manage topics."