#Preference-learning
Showing 5 of 5 repositories tagged #preference-learning, ranked by stars
Goekdeniz-Guelmez
MLX-LoRA-Studio
A native Mac App for LLM fine-tuning on Apple Silicon — fully on-device, fully open source.
Score
0
★ 254
⑂ 27
+3/day
Swift
IAAR-Shanghai
ICSFSurvey
Explore concepts like Self-Correct, Self-Refine, Self-Improve, Self-Contradict, Self-Play, and Self-Knowledge, alongside o1-like reasoning elevation🍓 and hallucination alleviation🍄.
Score
67
★ 173
⑂ 5
—
Jupyter Notebook
liushunyu
awesome-direct-preference-optimization
A Survey of Direct Preference Optimization (DPO)
Score
33
★ 94
⑂ 0
—
dengxianghua888-ops
ecoalign-forge
Multi-Agent DPO Data Synthesis Factory — 多智能体偏好训练数据自动合成框架 | 红队攻击 → 多persona审核 → 终审裁决 → DPO偏好对
Score
100
★ 71
⑂ 8
—
Python
JanoschMenke
metis
Python-based GUI to collect Feedback of Chemist in Molecules
Score
0
★ 54
⑂ 13
—
Python