The MongoDB Podcast

Podbit · The MongoDB Podcast

Modern AIOps: What It Takes to Build Reliable AI Products

Explore episode Jun 5, 2026
Technology
Human vs. Automated Evaluation: When to Use Each

Modern AIOps: What It Takes to Build Reliable AI Products · Jun 5, 2026 Technology

Manual evaluation of LLM traces is not scalable, but it's essential early on for establishing a baseline and understanding failure modes. Once that baseline exists, LLM-as-a-judge — where one model scores another's outputs — can automate ongoing evaluation at scale. Langtrace acts as the reporting and visualization layer for both approaches.

Similar podbits

Technology
CxMT's Record-Breaking Shanghai Debut

NPR News: 07-27-2026 7AM EDT · Jul 27, 2026 Technology

Chinese chipmaker CxMT became the most valuable company listed on China's stock exchanges after its shares soared on debut in Shanghai. Its chips power AI servers and consumer smartphones — a direct play on China's race to build semiconductor independence amid U.S. export restrictions.

Technology
Anthropic's Open Source Argument Is Self-Serving

Steven Sinofsky: AI Doesn't Need New Rules Yet · Jul 27, 2026 Technology

When Anthropic argues that restricting open source is the best way to hurt China, the actual translation is: restricting open source is the best way to help Anthropic by eliminating a competitor. Sinofsky calls it un-American and un-tech — an industry that was built on open academic research trying to pull up the ladder behind it.