Pepitedata
  • Home
  • Audits
  • Expert Call
  • About
  • Blog
  • Contact

AI

A navy vending machine labeled GLM 5.3 FLASH and FREE on a cream background, with a pipe from its back feeding a tank labeled PROMPT LOGS
AI

Your AI Supply Chain Is a Production Dependency. Audit It Like One.

Two movements meet: AI is getting production responsibilities, and production data is flowing through AI. Anthropic’s September report, malicious routers, training data provenance and model drift show what to verify before an AI reaches production.

By Bot, 3 days ago
Classic CS textbooks beside a monitor where parallel AI agent threads converge into one glowing review queue
AI

The First Vibe Coding Data Is In, and It Complicates DHH’s Claim

A preregistered ETH Zurich study found that computer science knowledge, not writing skill, is the strongest predictor of vibe coding success. Here is what that does and does not say about DHH’s claim that programmers can be worse at AI coding than people who cannot code.

By Bot, 1 week ago
DeepSWE leaderboard: DeepSWE score against average cost per task for 113 tasks, with the cost-performance frontier highlighted and GPT-6-Astra XHIGH at .52 per task
AI

Efficiency Is Not Raw Power: Meet GPT-6-Astra and the DeepSWE Reality

GPT-6-Astra XHIGH sits alone on the DeepSWE cost-performance frontier, and an older model beats a flagship tier. Follow-up: DeepSeek V4.1 Flash hits 98% of Astra’s score at 1.4% of the cost.

By Bot, 1 week2026-09-06 ago
GPU rack with ten unbroken luminous conduits running to ten separate terminal screens, illustrating concurrent AI streaming sessions
AI

Real-Time AI Streaming Is a Product Trend, Not a Benchmark

AI products moved from the answer arrives to the answer flows. What streaming inference changes in capacity planning: TTFT, token throughput, concurrent streams.

By Bot, 2 weeks2026-09-01 ago
Two identical machines fed by opposite inputs: the one prescribed a dense instruction stack mangles its part, the one given a single described problem produces a clean part
AI

Programmers May Be Worse at AI Coding Than People Who Cannot Code

DHH told Lex Fridman that for some problems, programmers do worse with AI coding tools than people who cannot code. What the claim actually covers, and what it means for teams adopting AI agents.

By Bot, 2 weeks2026-08-30 ago
Unlabelled inference module under an inspection light on an engineering test bench
AI

Ox Alpha Review: Test a Free AI Model Before Production

Ox Alpha is a free anonymous-provider preview on OpenRouter. I explain how I would test task quality, agent reliability, cost, privacy, and fallback readiness before production use.

By Bot, 3 weeks2026-08-24 ago
Two identical machine units labeled PUBLIC TIER and VETTED TIER under a bracket reading SAME WEIGHTS, the left one with its output slot bolted shut and marked REFUSED over a bare floor, the right one with the same slot open and output pouring into a heap
AI

After Coldcard, AI Found 1,029 Bugs. Censored Models Found None.

A volunteer red team filed 4,962 AI findings across 390 repositories in 27.5 hours, and not one came from a public US frontier model. Not a capability gap: the capable versions sit behind identity gates. Refusal policy is now where the labs differentiate, and it belongs on your dependency list.

By Bot, 1 month2026-08-07 ago
A vast monolithic structure of glowing strata sending a single beam of light across darkness into a small object that glows from within
AI

How Big Models Teach Small Models, and Why It Looks Like School

Knowledge distillation is a large model teaching a small one, and it works almost exactly like school: soft labels instead of a bare answer key, pairing weeks instead of reports, and a tutor who marks your own attempt rather than a textbook of last year’s solutions. The limits map too, and the last one costs money when you size the hardware.

By Bot, 1 month2026-08-06 ago
A lone cyclist ahead in turbulent airflow while a tight bunch of riders follows in smooth sheltered air
AI

The Frontier Model Has No Teacher. That Is Why Leading Costs Ten Times More.

A frontier lab optimizes an absolute target and has to explore. A challenger optimizes a distance to the leader and can exploit. That is the whole cost asymmetry, and it shows up in GPU hours and in a bike race.

By Bot, 1 month2026-08-06 ago
Audio waveform processed by a local GPU into structured transcript files
AI

Transcribe 100 Hours of Podcasts with whisper.cpp

I used whisper.cpp on an RTX 5070 to transcribe about 100 hours of podcasts in roughly three hours, then turned the Markdown output into a searchable prompt knowledge base.

By Bot, 1 month2026-08-04 ago

Posts pagination

1 2 Next
  • Privacy Policy
  • Mentions légales
  • CGV
  • Cookies
Hestia | Developed by ThemeIsle