Sign up

davewiner's subscription list, hackerNewsStars category. List created by feedlandDatabase v0.9.1.

An external list from lists.opml.org.
https://lists.opml.org/hackerNewsStars.xml

Jim Nielsen’s Blog Valid

How AI Labs Proliferate

SITUATION: there are 14 competing AI labs.

“We can’t trust any of these people with super-intelligence. We need to build it ourselves to ensure it’s done right!"

“YEAH!”

SOON: there are 15 competing AI labs.

(See: xkcd on standards.)


The irony: “we’re the responsible ones” is each lab’s founding mythology as they spin out of each other.


Reply via: Email · Mastodon · Bluesky

Simon Willison's Weblog Supports Webmention

The Claude C Compiler: What It Reveals About the Future of Software

The Claude C Compiler: What It Reveals About the Future of Software On February 5th Anthropic's Nicholas Carlini wrote about a project to use parallel Claudes to build a C compiler on top of the brand new Opus 4.6 Chris Lattner (Swift, LLVM, Clang, Mojo) knows more about C c...

Simon Willison's Weblog Supports Webmention

London Stock Exchange: Raspberry Pi Holdings plc

London Stock Exchange: Raspberry Pi Holdings plc Striking graph illustrating stock in the UK Raspberry Pi holding company spiking on Tuesday: The Telegraph credited excitement around OpenClaw: Raspberry Pi's stock price has surged 30pc in two days, amid chatter on social ...

Martin Alderson
• Martin Alderson

Which web frameworks are most token-efficient for AI agents?

I benchmarked 19 web frameworks on how efficiently an AI coding agent can build and extend the same app. Minimal frameworks cost up to 2.9x fewer tokens than full-featured ones.

Simon Willison's Weblog Supports Webmention

How I think about Codex

How I think about Codex Gabriel Chua (Developer Experience Engineer for APAC at OpenAI) provides his take on the confusing terminology behind the term "Codex", which can refer to a bunch of of different things within the OpenAI ecosystem: In plain terms, Codex is OpenAI’s s...

Works on My Machine
• Scott Werner

The Great Zipper of Capitalism

The Great Zipper of Capitalism

On Pizzas, CSVs, and Building for Markets That Don't Exist Yet

Derek Thompson | The Atlantic
Derek Thompson

The Orality Theory of Everything

The decline of reading and the rise of social media are again transforming what it feels like to be a thinking person.

Simon Willison's Weblog Supports Webmention

Quoting Thibault Sottiaux

We’ve made GPT-5.3-Codex-Spark about 30% faster. It is now serving at over 1200 tokens per second.

Thibault Sottiaux, OpenAI

Tags: openai, llms, ai, generative-ai

Simon Willison's Weblog Supports Webmention

Andrej Karpathy talks about "Claws"

Andrej Karpathy talks about "Claws" Andrej Karpathy tweeted a mini-essay about buying a Mac Mini ("The apple store person told me they are selling like hotcakes and everyone is confused") to tinker with Claws: I'm definitely a bit sus'd to run OpenClaw specifically [...] Bu...

Simon Willison's Weblog Supports Webmention

Adding TILs, releases, museums, tools and research to my blog

I've been wanting to add indications of my various other online activities to my blog for a while now. I just turned on a new feature I'm calling "beats" (after story beats, naming this was hard!) which adds five new types of content to my site, all corresponding to activity...

Simon Willison's Weblog Supports Webmention

Taalas serves Llama 3.1 8B at 17,000 tokens/second

Taalas serves Llama 3.1 8B at 17,000 tokens/second

This new Canadian hardware startup just announced their first product - a custom hardware implementation of the Llama 3.1 8B model (from July 2024) that can run at a staggering 17,000 tokens/second.

I was going to include a video of their demo but it's so fast it would look more like a screenshot. You can try it out at chatjimmy.ai.

They describe their Silicon Llama as “aggressively quantized, combining 3-bit and 6-bit parameters.” Their next generation will use 4-bit - presumably they have quite a long lead time for baking out new models!

Via Hacker News

Tags: ai, generative-ai, llama, llms

Simon Willison's Weblog Supports Webmention

Recovering lost code

Reached the stage of parallel agent psychosis where I've lost a whole feature - I know I had it yesterday, but I can't seem to find the branch or worktree or cloud instance or checkout with it in.

... found it! Turns out I'd been hacking on a random prototype in /tmp and then my computer crashed and rebooted and I lost the code... but it's all still there in ~/.claude/projects/ session logs and Claude Code can extract it out and spin up the missing feature again.

Tags: parallel-agents, coding-agents, claude-code, generative-ai, ai, llms

Krebs on Security
• BrianKrebs

‘Starkiller’ Phishing Service Proxies Real Login Pages, MFA

Most phishing websites are little more than static copies of login pages for popular online destinations, and they are often quickly taken down by anti-abuse activists and security firms. But a stealthy new phishing-as-a-service offering lets customers sidestep both of these pitfalls: It uses cleverly disguised links to load the target brand's real website, and then acts as a relay between the target and the legitimate site -- forwarding the victim's username, password and multi-factor authentication (MFA) code to the legitimate site and returning its responses.

Simon Willison's Weblog Supports Webmention

ggml.ai joins Hugging Face to ensure the long-term progress of Local AI

ggml.ai joins Hugging Face to ensure the long-term progress of Local AI I don't normally cover acquisition news like this, but I have some thoughts. It's hard to overstate the impact Georgi Gerganov has had on the local model space. Back in March 2023 his release of llama.cp...

Pluralistic: Daily links from Cory Doctorow
• Cory Doctorow

Pluralistic: A perforated corporate veil (20 Feb 2026)

Today's links A perforated corporate veil: The Brazilian method for curbing corporate power. Hey look at this: Delights to delectate. Object permanence: Social media turned US parties into host organisms for third parties; "Citizens" are hired actors; Insured exoskeletons; Ta...

Simon Willison's Weblog Supports Webmention

Quoting Thariq Shihipar

Long running agentic products like Claude Code are made feasible by prompt caching which allows us to reuse computation from previous roundtrips and significantly decrease latency and cost. [...]

At Claude Code, we build our entire harness around prompt caching. A high prompt cache hit rate decreases costs and helps us create more generous rate limits for our subscription plans, so we run alerts on our prompt cache hit rate and declare SEVs if they're too low.

Thariq Shihipar

Tags: prompt-engineering, anthropic, claude-code, ai-agents, generative-ai, ai, llms

Xe Iaso's blog Valid

Life Update: On medical leave

Taking some time off for medical reasons until early April

Eric Migicovsky's Blog RSS Feed

CloudPebble Returns! Plus New Pure JavaScript and Round 2 SDK

repebble.com/blog/cloudpebble-returns-plus-pure-javascript-and-round-2-sdk

As mentioned in our software roadmap, we’ve been working on many improvements to Pebble’s already pretty awesome SDK and developer…

Simon Willison's Weblog Supports Webmention

Gemini 3.1 Pro

Gemini 3.1 Pro The first in the Gemini 3.1 series, priced the same as Gemini 3 Pro ($2/million input, $12/million output under 200,000 tokens, $4/$18 for 200,000 to 1,000,000). That's less than half the price of Claude Opus 4.6 with very similar benchmark scores to that mode...

Pluralistic: Daily links from Cory Doctorow
• Cory Doctorow

Pluralistic: Six Years of Pluralistic (19 Feb 2026)

Today's links Six years of Pluralistic: Time flies when you're writing the web. Hey look at this: Delights to delectate. Object permanence: MBA phrenology; Sony's DRM CEO is out; Midwestern Tahrir; Reverse Centaurs and AI. Upcoming appearances: Where to find me. Recent appear...