The Hallway Track

fine-tuning

33 tracked signals on fine-tuning.

Frontier Tuning: Microsoft Build 2026

Microsoft Developer (Build) · Jun 03, 2026

Microsoft launches Frontier Tuning, letting enterprises reinforcement-fine-tune AI models on their own M365 data and workflows.

“With Frontier Tuning, we're making it possible for you to create your own enterprise AI.”
Announcing Foundry Managed Compute: Run open models in Microsoft Foundry

Microsoft AI Foundry Blog · Jun 03, 2026

Microsoft launches Foundry Managed Compute to run and customize open-source AI models on managed elastic GPU capacity.

“Managed Compute is built to lift that work off your team so the model, not the infrastructure, is what you spend your time on.”
How to go from your agent's traces to a fine-tuned model in one workflow

LangChain · Sep 24, 2026

LangChain launches LangSmith fine-tuning in public beta with SmithTune, a CLI to post-train models from agent traces.

“today we're launching LangSmith fine-tuning in public beta with SmithTune, a CLI to allow you to post-train models from your LangSmith traces in one workflow”
Frontier Tuning: Teaching AI to work the way you do

Microsoft 365 Dev Blog · Jun 02, 2026

Microsoft launches Frontier Tuning: RL-based AI customization within enterprise compliance boundaries

“a new approach to making AI work the way your business does by applying reinforcement learning inside your compliance boundary with your own data, processes, and conventions”
Stop writing longer prompts. Do this instead! 🧠

Google Developers (Google I/O) · Jun 24, 2026

Use dataset distillation and fine-tuning instead of longer prompts to enforce consistent structured outputs.

“Prompting tells the model what you want right now. Fine-tuning teaches the model a pattern it can follow repeatedly.”
Are Open Source Models Actually Ready for Production? | Spill The Tea

LangChain · Jun 13, 2026

Open source models still trail closed models on general agentic tasks but can outperform them once fine-tuned for a specific domain.

“because they're open and because you own the weights, you can actually fine-tune or RL them on your specific domain, which can actually make them perform better than those same closed models”
What Lies Beneath the API — Benjamin Cowen, Modal

AI Engineer · Jun 02, 2026

As AI products mature, more companies turn to fine-tuning over frontier APIs for performance and cost gains.

“If you tell your LLM to speak like a caveman, you can reduce your tokens by like a lot.”
Fine-tune a search agent with multi-turn RL on Amazon SageMaker AI

AWS Machine Learning Blog · Oct 02, 2026

Multi-turn RL on SageMaker lets small models match frontier reliability for search agents

“Fine-tuning offers a third path: you teach a small model your tools and environment directly. The result is a small model's speed and cost with the reliability that would otherwise require a frontier model.”
Build Small Hackathon

Hugging Face · Jun 05, 2026

Hugging Face kicks off the 'Build Small' hackathon celebrating small, fine-tunable models over large API providers.

Shipping custom models at scale from fine-tuning to inference | BRK234

Microsoft Developer (Build) · Jun 04, 2026

Fireworks AI's managed inference and training service is now available on Microsoft Azure.

“Fireworks is a managed service where we do both AI inference performance as well as AI training all as a managed service that you can now use on Microsoft Azure.”
Improve your agent’s tool-calling accuracy with SFT and DPO on Amazon SageMaker AI

AWS Machine Learning Blog · Jun 03, 2026

Combining SFT and DPO on SageMaker AI improves a small model's tool-calling accuracy for agents.

“When an agent picks the wrong tool, formats parameters incorrectly, or breaks a workflow chain, task completion times grow, error rates rise, support costs increase, and user experiences degrade.”
Training Agents 3: Reinforcement Learning

Hugging Face · Jul 28, 2026

Hugging Face is teaching GRPO reinforcement learning for agent training in a live stream series

“it's pretty straightforward to learn from the available options and you can apply it on most use cases”