<?xml version="1.0" encoding="UTF-8"?><rss version="2.0"><channel><title>Rohit Singh · Agentic AI Systems</title><description>Essays and posts on production agentic AI: multimodal RAG, self-healing infrastructure, and autonomous agents.</description><link>https://rohitsingh.ai/</link><item><title>A Second Check, Always</title><link>https://rohitsingh.ai/essays/a-second-check-always/</link><guid isPermaLink="true">https://rohitsingh.ai/essays/a-second-check-always/</guid><description>One model doing the work and an independent mechanism checking it is the single most reliable pattern I know for production AI. Here is how to build it.</description><pubDate>Wed, 02 Sep 2026 00:00:00 GMT</pubDate><category>essays</category><category>llm-as-judge</category><category>evaluation</category><category>reliability</category></item><item><title>Refuse Over Guess</title><link>https://rohitsingh.ai/essays/refuse-over-guess/</link><guid isPermaLink="true">https://rohitsingh.ai/essays/refuse-over-guess/</guid><description>Why the most important design decision in a production AI system is deciding when the model is not allowed to answer, and how to engineer that refusal.</description><pubDate>Wed, 02 Sep 2026 00:00:00 GMT</pubDate><category>essays</category><category>agentic-ai</category><category>reliability</category><category>verification</category></item><item><title>Kimi K3: The Architecture Was Decided By The Kernel</title><link>https://rohitsingh.ai/blog/kimi-k3-architecture-decided-by-the-kernel/</link><guid isPermaLink="true">https://rohitsingh.ai/blog/kimi-k3-architecture-decided-by-the-kernel/</guid><description>Reading the Kimi K3 technical report from the seat of someone who builds agentic systems that have to run on Monday morning: the decisions were made by kernels, caches, and harnesses, not loss curves.</description><pubDate>Thu, 30 Jul 2026 00:00:00 GMT</pubDate><category>model-analysis</category><category>agentic-ai</category><category>inference</category><category>evaluation</category></item><item><title>Skills, Connectors, and Subagents: Decoding the Architecture Anthropic Just Made Standard</title><link>https://rohitsingh.ai/blog/skills-connectors-subagents/</link><guid isPermaLink="true">https://rohitsingh.ai/blog/skills-connectors-subagents/</guid><description>Anthropic&apos;s agent templates for regulated industries matter less for the customer logos than for the orchestration topology underneath: lazy-loaded skills, governed connectors, and isolated subagents.</description><pubDate>Sat, 09 May 2026 00:00:00 GMT</pubDate><category>model-analysis</category><category>agentic-ai</category><category>architecture</category><category>regulated-industries</category></item><item><title>The Open-Source Agentic AI: Kimi K2.6</title><link>https://rohitsingh.ai/blog/open-source-agentic-ai-kimi-k2-6/</link><guid isPermaLink="true">https://rohitsingh.ai/blog/open-source-agentic-ai-kimi-k2-6/</guid><description>An honest read of Moonshot&apos;s Kimi K2.6 from someone building production agentic systems: why the open weights matter more than the leaderboard, and where it still falls short.</description><pubDate>Wed, 22 Apr 2026 00:00:00 GMT</pubDate><category>model-analysis</category><category>agentic-ai</category><category>open-source</category><category>evaluation</category></item><item><title>Building Smarter AI Benchmarks</title><link>https://rohitsingh.ai/blog/building-smarter-ai-benchmarks/</link><guid isPermaLink="true">https://rohitsingh.ai/blog/building-smarter-ai-benchmarks/</guid><description>What two controversial papers, The Illusion of Thinking and its rebuttal, taught us about measuring machine reasoning: many AI failures are benchmark design failures.</description><pubDate>Thu, 19 Jun 2025 00:00:00 GMT</pubDate><category>evaluation</category><category>benchmarks</category><category>reasoning</category></item><item><title>HITL vs Full Automation: A Decision Framework</title><link>https://rohitsingh.ai/blog/hitl-vs-full-automation/</link><guid isPermaLink="true">https://rohitsingh.ai/blog/hitl-vs-full-automation/</guid><description>When to keep humans in the loop and when to ship fully autonomous agentic systems.</description><pubDate>Wed, 20 Mar 2024 00:00:00 GMT</pubDate><category>agentic-ai</category><category>hitl</category><category>automation</category><category>decision-framework</category></item><item><title>5 RAG Failure Modes in Production</title><link>https://rohitsingh.ai/blog/rag-failure-modes/</link><guid isPermaLink="true">https://rohitsingh.ai/blog/rag-failure-modes/</guid><description>Common ways RAG systems fail in week 2, and how to avoid them with evaluation and observability.</description><pubDate>Fri, 15 Mar 2024 00:00:00 GMT</pubDate><category>rag</category><category>production</category><category>evaluation</category><category>langfuse</category></item></channel></rss>