Blog Archive
X as Code Is Not a File Conversion: What Must Survive When Engineering Leaves Its Tool?
Architecture may be stored in Rhapsody, requirements in DOORS or Codebeamer, decisions in Word, and interfaces in ARXML. Moving that system to Git is not a serialization problem. It is a decision about which engineering meaning must survive.
Read Post
Why I’m Building SWCraft: Keeping the Architect in Control of AI-Assisted AUTOSAR Design
SWCraft explores how deeply AI can participate in AUTOSAR architecture development without taking control of the engineering decision.
Read Post
My MCP Server Had 39 Tools. I Had Designed It for the Wrong Caller.
SWCraft grew to 39 MCP tools by exposing model operations one tool at a time. Separating the software API from the agent interface reduced the surface to 13 without removing capability.
Read Post
An LLM Wrote the Code. ISO 26262 Doesn't Care — and That's the Point.
ISO 26262 tool qualification never asks who wrote the code, because it never trusted the tool in the first place. If the confidence argument does not depend on the AI being right, AI-assisted development and qualification stop being in conflict.
Read Post
Nothing Changed in the API. Everything Changed for the Agent.
An MCP tool schema explains how to call a function, but not when that function should become a candidate. A task-scoped semantic contract can keep a growing tool catalog behaviorally compatible.
Read Post
An Agent Can Pass Validation and Still Make an Unsafe Change
The agent was authorized to modify one component. It helpfully modified two. The model still validated. Why MCP capability is not task authority, and how a task-scoped mutation boundary closes the gap.
Read Post
The Requirements RAG Router Failed. Then the Benchmark Failed Too.
A requirements RAG reliability experiment: heuristic routing failed on independent paraphrases, query grounding exposed a broken benchmark taxonomy, and typed evidence plans produced a narrow but useful result.
Read Post
Everyone Is Building an AI AUTOSAR Toolchain. But How Should We Evaluate One?
Generating plausible AUTOSAR artifacts is an interesting capability. It is not yet evidence of a production-ready engineering system. We need a way to measure the difference.
Read Post