Luke Angel
← back to the journal
Tag

#dgx-spark

10 entries
tools
FLUX Can Spell, SDXL Cannot: Local AI Image Models on a DGX Spark
“A caption written after looking will always find something to praise.”
Aug 21
method
Forty-Four Harness Bugs, Zero Local LLM Limitations: an Accounting
“A model that looks broken is a harness bug until proven otherwise.”
Aug 19
tools
Nine Local LLMs Ranked by Cost Per Solved Task — Seven of Them Have No Price at All
“A repair that doesn't satisfy the contract isn't 80% of a repair. It's a workspace you throw away.”
Aug 17
method
What a Bug Fix Costs on Two DGX Sparks: 16 Concurrent AI Agents, 4.45× Cheaper Than Serial
“The ceiling I'd been running against was a benchmark's sample size, not a limit of the machine.”
Aug 14
method
Nine Lines of Verification That Beat a Six-Agent AI Swarm
“It tried to stop 52 times. It was refused 51. That is the whole mechanism.”
Aug 12
method
1,200 Lines of Multi-Agent Orchestration, Beaten by One Local LLM Agent
“The orchestration didn't lose because multi-agent is a bad idea. It lost because it was insurance against a blindness that no longer existed.”
Aug 10
craft
The Day I Lost to Tensor Parallelism: Nemotron-70B Across Two DGX Sparks
“The right response to 'the weights don't fit' was to shrink the weights, not to distribute them across a fabric with a documented collective bug.”
Aug 09
tools
The Agent Framework Bake-Off: LangGraph vs Pydantic AI vs Hand-Rolled, and the 32 Lines That Mattered
“The gain didn't come from primitives. It came from 32 lines being harder to get wrong.”
Aug 07
method
Twenty-Two Behavioural Criteria: What Replaced the Structural Grader My AI Agents Kept Fooling
“Scoring went from 28 seconds to 2, and got strictly harder to fool.”
Aug 06
method
The AI Grader That Passed the Same Bug Three Times: 229 Checks, Three Agent Workspaces, One Blind Spot
“A perfect score, three times, over the same bug. The checks weren't wrong. They were pointed at the wrong thing.”
Aug 04