Preprint2024
LearnLM
Unknown
LearnLM enhances Gemini for pedagogical instruction by co-training SFT and RLHF, outperforming leading LLMs in tutoring scenarios.
0Dec 1, 2024Fine TuningReinforcement Learning
arXiv
A comprehensive index of artificial intelligence and machine-learning research with AI-generated summaries, citation metrics, and direct links to papers and code.
Unknown
LearnLM enhances Gemini for pedagogical instruction by co-training SFT and RLHF, outperforming leading LLMs in tutoring scenarios.
Luzhe Huang, Hanlong Chen, Tairan Liu, et al.
GedankenNet eliminates the need for labeled or experimental training data in hologram reconstruction by using a physics-consistency loss and synthetic random images.
Unknown
This paper introduces Orfs-agent, a tool-using reinforcement learning agent that optimizes chip design workflows by selecting and sequencing EDA tools.