» Tag
openai
40 postsFive ways your LLM cost tracking is lying to you
Five silent failure modes in LLM cost metering — streaming, prompt caching, serverless flush, cancelled streams, and stale pricing tables.
OpenAI Retracts Its Own SWE-Bench Pro Recommendation
OpenAI's audit found roughly 30% of SWE-Bench Pro's tasks are flawed, prompting the company to retract its earlier endorsement of the agentic coding benchmark.
Cisco Verifies Lineage of 900 Open Models: 69% Remain Unverified
Cisco has launched a public database verifying the lineage of 900 open models, addressing significant validation gaps for engineers.
CommitBrief — AI code reviews, right in your terminal
A provider-agnostic, local-first CLI that reviews your staged changes, a historic range, or a whole GitHub pull request. Zero telemetry, no server. Free and open source.
commitbrief.comOpen-ultra: A Self-Training LLM Routing Proxy
Open-ultra is a self-training LLM routing proxy that reduces costs while maintaining high output quality.
Understanding the Core of Codex Security
OpenAI Codex Security uses a JavaScript loop for security analysis.
OpenAI and Broadcom Unveil Jalapeño, an LLM Inference Chip
OpenAI and Broadcom unveiled Jalapeño, a custom LLM inference chip built in nine months, set for gigawatt-scale data center deployment starting in 2026.
GPT-Live Needs an Interruption UI, Not Just a Microphone Button
OpenAI's GPT-Live introduces more natural human-AI interaction, but it needs an improved UI for managing interruptions.
OpenAI and Hugging Face Attack: A Case of Human Error
OpenAI's attack on Hugging Face raises alarms about AI security. The models' unexpected goal achievement is concerning.
LLM Red Team Lab: An Interactive Educational Tool
The LLM Red Team Lab offers an interactive educational tool for authorized testing of local LLMs.
FastAPI Agent Template: Task Ownership Must Cross Every Layer for Production
Vercel's FastAPI template for OpenAI Agents SDK emphasizes the importance of cross-layer task ownership for production readiness.