Claude & Anthropic
34diegosouzapw/OmniRoute: Never stop coding. Free MIT AI gateway: one endpoint, 268+ providers (50+ free), 500+ models — Kimi, Claude, GPT, OpenAI, Gemini, GLM, DeepSeek, MiniMax. Works with Claude Code, Codex, Cursor, OpenCode, Cline & Copilot. Quota-aware auto-fallback, RTK+Caveman compression saves 15-95% tokens, MCP/A2A, Desktop/PWA. Built by 500+ contributors
Judge approves $1.5B Anthropic settlement for pirated books used to train Claude
Announcing Ship: an endpoint with the highest intelligence per dollar of any frontier model. Today, Ship makes using Opus and GPT 5.6 Sol 50% cheaper
@@withmartian
Don't upgrade to the $100 Claude plan (yet). These 21 hacks make the $20/month plan enough: 1. You upload PDFs raw. One page = 3,000 tokens. Fix: P
@@sauda_coder
In 1948, Claude Shannon invented the math behind every LLM you use today. He tested it by making his wife guess the next letter in a book. A Stanford
@@_yusufknl
Anthropic just released Claude Sonnet 5 — and it changes the calculation for anyone building agents. The short version: Sonnet 5 gets 63.2% on agenti
@@TeksCreate
I Made Claude Code and Codex Argue About My Code Until They Agreed
TL;DR: I wired OpenAI's Codex CLI into Claude Code as an adversarial reviewer with a convergence...
How Anthropic runs large-scale code migrations with Claude Code
Kimi K3 vs Claude Opus 4.8: Benchmarks, Price, Verdict
Kimi K3 ties Claude Opus 4.8 on GPQA Diamond, costs 40% less per token, and its weights are expected by July 27 — but Opus still leads where it counts for some teams. A fact-checked comparison of benc
What survives compaction in Claude Code — and how to keep your rules alive
Long Claude Code sessions eventually hit the context window ceiling. When they do, Claude Code runs...
How to Use Kimi K3 with Claude Code, Cursor, and Cline
Kimi K3 in Claude Code takes three environment variables. This guide walks through the exact setup for Claude Code, Cursor, and Cline via LLM Gateway — plus what each tool does and doesn't route, and
Agentic & Tools
23Agents in the Wild: Where Research Meets Deployment
Agentic systems large language model (LLM) based architectures capable of reasoning, planning, acting, and coordinating with tools and other agents are rapidly transitioning from research prototypes t
Graph-Based Agentic AI with LangGraph: Workflow Pathways for Long-Running Stateful Business Processes
This paper is a practitioner guide to graph-based workflow pathways for long-running, stateful, multi-step generative AI systems in business processes. Rather than treating LangGraph, a low-level orch
KnockOutEZ/wigolo: The go-to web for your AI coding agent — local-first search, fetch, crawl & research over MCP. No API keys, no cloud, $0/query. Public beta.
FilmWorld: Agentic Novel-to-Film Generation
tirth8205/code-review-graph: Local-first code intelligence graph for MCP and CLI. Builds a persistent map of your codebase so AI coding tools read only what matters, with benchmarked context reductions on reviews and large-repo workflows.
Show HN: MindCache – An open-source agentic memory system for LLMs
Choosing GPT-5.6 Sol, Terra, or Luna in Codex
Tool Schema Drift: The Silent Failure Mode in Production Agentic Systems
The most common agentic system failure I encounter in production is not a bad prompt. It is not a...
They'll Verify. They Just Won't Act. How Authority Framing and Laundered Code Turn a Trusted Agentic CI/CD Pipeline Into an Attack Surface
We study a five-agent CI/CD pipeline (triage -> developer -> security-scan -> review -> approve/deploy), built from five distinct production LLMs across three providers, behind an LLM firewall in shad
Toward Auditable Fraud Detection: Combining Graph Features, Model Explanations, and Agentic Case Investigation
Fraud detection systems must scale with rising transaction volume while remaining explainable and reviewable. We study a layered pipeline on the PaySim dataset that combines a gradient-boosted classif
AI = system{model weights, model config, harness, datasets, hardware optimization, product UI/UX, sandboxes, other agentic engineering...}. Modern d
@@varun_mathur
Today we're launching TokenSwitch. For the past year, CTOs were giving their developers whichever coding agent they preferred and letting them token
@@ashtilawat
Models & Releases
11Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber
Kimi K3 Is Competitive with Fable; Kimi K3 and Fable Is SoTA
To summarize: HuggingFace got autonomously compromised by a model from an American company. HF then tried to use American frontier model(s) to defend
@@attrc
Gemini last models: temperature, top_p, and top_k are deprecated and ignored
You should be partnering with @huggingface to release the model weights you have deprecated #OpenSource4o #OpenSourceo3 #OpenSource41 #OpenSource45
@@yv_thorne
unclecode/crawl4ai: A unified library of SOTA model optimization techniques like quantization, distillation, pruning, neural architecture search, speculative decoding, etc. It compresses deep learning models for downstream deployment frameworks like TensorRT-LLM, TensorRT, vLLM, etc. to optimize inference speed.
Gemini 3.6 Flash released on AIStudio
Google releases three new Gemini models — but no 3.5 Pro
Google released Gemini 3.6 Flash, 3.5 Flash-Lite, and Flash Cyber, but the continued absence of Gemini 3.5 Pro raises fresh questions about its AI strategy.
Google announces Gemini 3.6 Flash and cybersecurity AI, teases 3.5 Pro and Gemini 4
There are new 3.6 and 3.5 models today, but Google is already training Gemini 4.
New Gemini models dropped https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-3-6-flash-3-5-flash-lite-3-5-flash-cyber/
3.6...
underrated part of the story is huggingface couldn't use oai or ant models to investigate the incident due to guardrails, and were forced to use chine
@@shanejcaldwell
Industry & General
7This was our first incident of this kind, and we want to thank OpenAI for its transparency about what happened and for the collaboration. Fortunately
@@Thom_Wolf
Microsoft AI CEO Mustafa Suleyman explains the exit plan behind the biggest partnership in AI. It began with a Satya Nadella message surfaced in the t
@@karlmehta
BREAKING $AMD 🤝Cornelis on EPYC Venice & MI400🚀🚨🆕 Cornelis Announces New Reference Architecture for AI Inference, Training, and HPC Built for AMD 6th
@@MikeLongTerm
Show HN: A self-running space economy SIM in Rust and Bevy
This was the most packed week in AI I've ever seen. 15 major updates in 7 days. If you missed even one day, you're behind. Here's everything that hap
@@VaibhavSisinty