The basic idea here is just automating and limiting the steps that AI can take. These are described as ‘nodes’ and their actions are limited/more descriptive from developer perspective.
The langgraph simply allows you to set some fences around the AI agent for a goal, instead of raw terminal flow. Is it better? Arguable.
Comparison is simply byte by byte equalness check of reference kernel outputs with candidate (optimized) outputs.
Why I added ai generated inputs then? I was just being lazy and this was more of a langgraph playground for me:)