
LangSmith introduces a self-improving LLM-as-a-Judge system that eliminates the need for extensive prompt engineering, allowing users to set up evaluators with minimal configuration. The system stores user corrections as few-shot examples, which are used to refine the evaluator's performance over time, enabling it to adapt to human preferences without manual intervention. This approach streamlines evaluation by automatically integrating real-world feedback into the model's training process.
Managed Deep Agents gives developers a managed way to build, run, and deploy Deep Agents with built-in runtime, streaming, sandboxes, evals, memory, and auth.

LangSmith Bring Your Own Cloud is now generally available on AWS, giving Enterprise teams managed observability, evaluation, and deployment inside their own VPC.

Learn what AI agents are, how they work in an LLM loop, and where workflows fit so you can build reliable, production-ready autonomous systems.
