<?xml version="1.0" encoding="UTF-8"?><rss version="2.0"><channel><title>Konstantin Tarandevich</title><description>Personal blog and portfolio of Konstantin Tarandevich, Software Architect at ZF Group based in Koblenz, Germany.</description><link>https://site.tarandevich.org/</link><item><title>Everyone Is Building an AI AUTOSAR Toolchain. But How Should We Evaluate One?</title><link>https://site.tarandevich.org/blog/everyone-is-building-an-ai-autosar-toolchain/</link><guid isPermaLink="true">https://site.tarandevich.org/blog/everyone-is-building-an-ai-autosar-toolchain/</guid><description>Generating plausible AUTOSAR artifacts is an interesting capability. It is not yet evidence of a production-ready engineering system. We need a way to measure the difference.</description><pubDate>Sun, 19 Jul 2026 00:00:00 GMT</pubDate></item><item><title>Hybrid Databases Were the Wrong Fix for Requirements RAG</title><link>https://site.tarandevich.org/blog/hybrid-databases-wrong-fix-requirements-rag/</link><guid isPermaLink="true">https://site.tarandevich.org/blog/hybrid-databases-wrong-fix-requirements-rag/</guid><description>A negative infrastructure result from a requirements RAG lab: Postgres, ParadeDB, and OpenSearch could host useful pieces, but none replaced the current retrieval stack without losing ranking quality or evidence bundles.</description><pubDate>Wed, 24 Jun 2026 00:00:00 GMT</pubDate></item><item><title>The RAG Verifier Learned to Say No. It Still Missed a Hard Requirement Failure.</title><link>https://site.tarandevich.org/blog/claim-level-verification-requirements-rag/</link><guid isPermaLink="true">https://site.tarandevich.org/blog/claim-level-verification-requirements-rag/</guid><description>A held-out requirements RAG experiment: claim-level verification reduced harmful published answers from 35.7% selective risk to 4.0%, but strict human audit still found semantic completeness failures.</description><pubDate>Sun, 21 Jun 2026 00:00:00 GMT</pubDate></item><item><title>Citations Were Not Enough for Safety-Critical Requirements RAG</title><link>https://site.tarandevich.org/blog/citations-were-not-enough-safety-critical-requirements-rag/</link><guid isPermaLink="true">https://site.tarandevich.org/blog/citations-were-not-enough-safety-critical-requirements-rag/</guid><description>A negative result from a requirements RAG lab: even with structured records, cited JSON answers, and deterministic guards, the answerer failed the abstention gate when evidence was insufficient.</description><pubDate>Mon, 15 Jun 2026 00:00:00 GMT</pubDate></item><item><title>I Flattened 13,244 Requirements Into Chunks. 97% Lost Their Meaning.</title><link>https://site.tarandevich.org/blog/rag-structured-requirements-router/</link><guid isPermaLink="true">https://site.tarandevich.org/blog/rag-structured-requirements-router/</guid><description>A practical retrieval study over structured automotive requirements: why flatten-and-chunk failed, where ordinary top-k search stopped being the right tool, and how query routing improved both quality and latency.</description><pubDate>Sun, 07 Jun 2026 00:00:00 GMT</pubDate></item><item><title>I Tested a Simple RAG Idea: One Summary per Document. It Was Useful, But Not Enough.</title><link>https://site.tarandevich.org/blog/rag-retrieval-lab-summary-retrieval/</link><guid isPermaLink="true">https://site.tarandevich.org/blog/rag-retrieval-lab-summary-retrieval/</guid><description>A practical retrieval experiment: can an LLM-generated document summary replace chunk-level search? On a 5,000-document corpus, the answer was measurable — useful signal, clear ceiling, and a better baseline.</description><pubDate>Sun, 31 May 2026 00:00:00 GMT</pubDate></item><item><title>My Two AI Agents Talk MCP to Each Other. There Is No Standalone Tool Server.</title><link>https://site.tarandevich.org/blog/agents-talking-mcp-without-a-tool-server/</link><guid isPermaLink="true">https://site.tarandevich.org/blog/agents-talking-mcp-without-a-tool-server/</guid><description>The default MCP story is one fat tool server, many clients. When I needed two of my domain agents to talk to each other, that shape was the wrong one — and the alternative shows what MCP is actually good for.</description><pubDate>Thu, 28 May 2026 00:00:00 GMT</pubDate></item><item><title>I Built an AI Agent for Myself. My Colleagues Wanted to Use It. That Is Where the Problems Started.</title><link>https://site.tarandevich.org/blog/ai-agent-bus-factor-of-one/</link><guid isPermaLink="true">https://site.tarandevich.org/blog/ai-agent-bus-factor-of-one/</guid><description>When a domain expert builds a working AI agent and people start depending on it, the bus factor problem arrives long before the platform — and a different kind of work begins.</description><pubDate>Thu, 14 May 2026 00:00:00 GMT</pubDate></item><item><title>The Boring Part of Requirements Review Is Now Automated</title><link>https://site.tarandevich.org/blog/ai-agent-requirements-review/</link><guid isPermaLink="true">https://site.tarandevich.org/blog/ai-agent-requirements-review/</guid><description>How a lightweight AI agent eliminates the context-gathering overhead in safety-critical requirements review — so the architect can focus on judgment, not tool-switching.</description><pubDate>Sat, 02 May 2026 00:00:00 GMT</pubDate></item><item><title>Measuring What Matters: Architecture Quality Metrics for Safety-Critical Software</title><link>https://site.tarandevich.org/blog/measuring-what-matters-architecture-quality-metrics/</link><guid isPermaLink="true">https://site.tarandevich.org/blog/measuring-what-matters-architecture-quality-metrics/</guid><description>Architectural quality is difficult to observe — you cannot run a test suite on a dependency graph. This post describes a metrics system built to track it: three levels of measurement, from binary compliance checks to long-term trends.</description><pubDate>Tue, 28 Apr 2026 00:00:00 GMT</pubDate></item><item><title>Building an AI Agent for Embedded Systems Architects</title><link>https://site.tarandevich.org/blog/building-ai-agent-for-embedded-systems-architects/</link><guid isPermaLink="true">https://site.tarandevich.org/blog/building-ai-agent-for-embedded-systems-architects/</guid><description>A conversational AI agent that sits in front of complex embedded systems tooling and speaks human — so architects spend less time remembering plugin syntax and more time on actual architecture.</description><pubDate>Mon, 27 Apr 2026 00:00:00 GMT</pubDate></item></channel></rss>