This week

What landed in the last seven days — labs, papers, releases.

llama.cpp b10976

ci : fix android release ( #28936 ) Website: https://llama.app Attestations: https://github.com/ggml-org/llama.cpp/attestations/47548995 macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFramework Linux: Ubuntu…

release

opt in/out

Via [AINews] AEF-1 standard emerges for Third Party Evaluators, as Xai, OpenAI, and Anthropic all cosign

news

rare personal blogpost

I have worked on AI for the last twelve years because I believe it could dramatically raise the quality of human life. I’ve written often about these incredible benefits: I believe that AI could cure most major diseases in the next 5–10 years, greatly accelerate economic growth…

tool

Show HN: Sunk Cost – How long until a local LLM rig pays for itself?

I kept hearing "just buy a Mac and run models locally, it pays for itself" and wanted to check. Sunk Cost takes a machine, a model and how many tokens you use a day, and works out how long the hardware takes to pay back against renting the same model by the token. Obviously…

news

Anthropic is in regulatory-capture financial loop

The AI safety nonprofits around Anthropic promote an Anthropic-aligned Doomer narrative, whether Tarbell Fellows in popular media, or METR in evaluations. All rely on Moskovitz's Anthropic stock worth $7 billion, which the parent funding org sits on. This stock increases in…

news

llama.cpp b10970

HIP: fattn-mma: use fp32 accumulation on MFMA devices ( #28576 ) use fp32 accumulators in fattn-mma on CDNA Website: https://llama.app Attestations: https://github.com/ggml-org/llama.cpp/attestations/47450174 macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64,…

release

llama.cpp v0.4.1

Overview llama.cpp 0.4.1 adds Maple 20B-A1B, Tencent Hy 4, and Spark2.5 support, improves JSON schema handling, chat parsing, logging, and server child-process management, and updates ggml to v0.24.0. API changes Changed llama_sampler_chain_n() to return int32_t instead of int (…

release