Open Source Radar — August 25: Karpathy's playbook, open-sourced

Share
Open Source Radar — August 25: Karpathy's playbook, open-sourced

Today's board has a single quiet fingerprint: Andrej Karpathy's recurring complaints about how coding agents misbehave have been turned into an instructions file people actually install, and his "LLM Wiki" idea now underpins a self-organizing note system. Around them sit the rest of the stack that thread depends on — a coding agent that runs in your terminal, and a router that makes free LLM inference worth stacking. Four fresh projects, none repeats from the week's earlier radars.


Andrej Karpathy Skills (config, ~206k stars) — The biggest mover on the daily board is just one instructions file, but it's the one people are installing everywhere. It takes Karpathy's running list of LLM coding pitfalls — making wrong assumptions silently, bloating code with dead abstractions, refactoring files they were never asked to touch — and compresses them into four working principles: think before coding, simplicity first, surgical changes, and goal-driven execution. Drop it into Claude Code, Cursor, or Codex and your agent is supposed to ask clarifying questions before guessing, keep diffs minimal, and chase verifiable success criteria instead of vague "make it work." At over 200k stars in days, it's the clearest signal yet that prompt discipline is the cheapest quality upgrade on the agent stack.


Codex (Rust, ~107k stars) — OpenAI's terminal coding agent hit daily trending hard, and it's the one that ties today's thread together because it's the kind of agent those discipline files are meant to tame. It runs locally on your machine, plans before it edits, and works whether you sign in with a ChatGPT plan or pass it an API key. It's been a fixture of the agent stack for a while, but the open-source, Apache-licensed CLI keeps drawing stars on the strength of being a serious, lightweight coding agent that doesn't lock you into a heavy harness. You'd reach for it when you want an assistant that can actually run code and verify results in your repo without standing up a whole platform around it.


FreeLLMAPI (TypeScript, ~19k stars) — An OpenAI-compatible proxy that stacks the free tiers of roughly 28 LLM providers — about 4 billion tokens a month across 358 free endpoints — behind a single local server. Point any OpenAI-style client at one address and it routes each request to the best available free model, fails over when a provider throttles you, and tracks per-key usage so you stay under every cap. It's honest about its limits: the maintainer says it's for personal experimentation and learning, not production, and you swap in a paid API before shipping anything real. For prototyping, tinkering, or running coding agents under a strict budget, it turns a pile of per-provider toys into something usable.


claude-obsidian (Python, ~11.5k stars) — A self-organizing "second brain" for Obsidian driven by Claude Code, built directly on Karpathy's LLM Wiki pattern. Drop in any source — a PDF, a webpage, meeting notes — and the agent reads it, links it, and files it into a connected knowledge graph of plain Markdown that stays yours, not locked in a plugin cache or a cloud database. It grounds claims in a source ledger, builds linked pages and Maps of Content, and answers later questions from what the vault already knows instead of starting every session from a blank slate. It's a local-first, MIT-licensed alternative to hosted note apps for anyone who wants their notes to keep compounding in value without giving up ownership of the files.

Worth watching this week: Karpathy's maxims have stopped being tweets and started being shipped tools — the agents people actually use are increasingly governed by files that say "don't assume, don't over-engineer, verify."

If your agent could file everything you read into a knowledge graph you own, would you trust it — or keep taking notes by hand? Tell us in the comments.

Sources: Andrej Karpathy Skills (GitHub) · Codex (GitHub) · FreeLLMAPI (GitHub) · claude-obsidian (GitHub)

Read more

OpenAI busts influence ops that planted fake stories in real media

OpenAI busts influence ops that planted fake stories in real media

The day's AI news runs through one seam: the work is showing up in places nobody planned for — inside real newsrooms, across the whole night sky, and in the M&A column. OpenAI has banned two state-backed influence operations that used ChatGPT to plant fabricated stories inside legitimate news outlets — and rated the Russian one the most disruptive it has seen in two and a half years. In a report dated October 8, OpenAI detailed "Dark Clark," run from Russia across Latin America, which ran a th

Open Source Radar — October 9: plugins, sandboxes, tokens

Open Source Radar — October 9: plugins, sandboxes, tokens

Today's open-source signal is infrastructure rather than hype: Microsoft's code sandbox reaches 1.0, Anthropic's knowledge-worker plugins keep climbing, a beloved token counter flips its default, and LocalLLaMA squeezes a usable 2B model into about 700 MB. knowledge-work-plugins (Python, ~27,900 stars, Apache-2.0) — Anthropic's repository of role-shaped plugins for Claude Cowork is the top AI repository on today's daily trending page, and the stars keep coming: roughly 2,100 more than when we

Deep Dive — The four-token blind spot inside DeepSeek V4

Deep Dive — The four-token blind spot inside DeepSeek V4

ByteDance's Seed research team says it has found the cause of one of the stranger recurring complaints about DeepSeek's models: the same question, asked with nothing changed except a few junk characters bolted onto the front, can flip the model from right to wrong. Their paper, posted to arXiv on September 28, traces the wobble to a memory-saving trick used during long-context inference, and reports that DeepSeek-V4-Flash-Base's retrieval accuracy swings by as much as 40.2 percentage points depe

SoftBank seeks $100B from Gulf investors for an AI fund

SoftBank seeks $100B from Gulf investors for an AI fund

Three moves today point the same direction: the money, the politics, and the price of speed all got more expensive. SoftBank is reportedly seeking up to $100 billion from Gulf investors for a fund that would buy companies and run them with AI. The Financial Times reported the raise, citing people familiar with the matter, and says Masayoshi Son has held discussions in recent weeks with senior figures including in the United Arab Emirates; Reuters and Bloomberg both carried the report but neith