about
Stories from July 29, 2025 (UTC)
Go back a day, month, or year. Go forward a day.
1. Supervised fine tuning on curated data is reinforcement learning (arxiv.org)
71 points by GabrielBianconi on Jul 29, 2025 | hide | past | pdf | 19 comments
2. Query Agnostic Adversarial Triggers for Reasoning Models (arxiv.org)
3 points by fzliu on Jul 29, 2025 | hide | past | pdf | discuss
3. Language Model Can Be a Steganographic Privacy Leaking Agent (arxiv.org)
3 points by dennis-tra on Jul 29, 2025 | hide | past | pdf | discuss
4. TrimLLM: Progressive Layer Dropping for Domain-Specific LLMs (arxiv.org)
2 points by pulkitsh1234 on Jul 29, 2025 | hide | past | pdf | discuss
5. SmallThinker: A Family of Efficient LLMs Natively Trained for Local Deployment (arxiv.org)
2 points by limoce on Jul 29, 2025 | hide | past | pdf | discuss