This page collects what the sources on blogs to watch have published recently and groups them by subject rather than by publisher. It is rebuilt every hour, so a post usually appears here within an hour of going up. Nothing on this page is required reading.
Grouping by subject is the point. A reader gives you 45 separate streams and leaves you to notice that four companies wrote about the same scheduling problem this week; this page puts those four posts under one heading. Topics come from the filter terms already listed on the blogs-to-watch page, so the vocabulary is the course's.
How a post is filed. Each post is scored by keyword against every topic, using its title and its summary, and it is filed under the topic it scores highest against. A term that names a subject on its own counts for more than one that merely co-occurs with it, and a match in the title counts for three times a match in the summary. Any second topic a post also matches is shown as a label beside it. The method is keyword matching rather than a model, which makes it predictable, cheap, and occasionally wrong.
You are reading one topic, everything else, over the last 120 days. Show every topic.
Posts from the watchlist that no topic above claims.
Learn how mortgage document automation transforms loan processing, from document ingestion and extraction to validation and system integration.
Learn how KYC automation replaces manual verification with scalable, compliant workflows that cut costs, reduce errors, and speed up onboarding.
Learn how unstructured data extraction turns documents, PDFs, and text into structured insights using AI, NLP, and LLMs for scalable data processing.
OCR for tables converts complex document layouts into structured, machine-readable data. Learn how LlamaParse preserves table integrity.
OCR for images helps convert photos, labels, and screenshots into structured text. Compare the top AI OCR tools and learn what makes a reliable image-to-text system.
Learn how OCR for accounts payable automates invoice processing, improves accuracy, reduces costs, and integrates structured data into ERP systems.
OCR that works in a demo often stalls in production. How modern document pipelines handle real corpora, and how to hit 90%+ straight-through rates.
Intelligent OCR turns documents into validated, structured data, not just text. How production pipelines handle parsing, extraction, and confidence.
LlamaIndex is a simple, flexible framework for building knowledge assistants using LLMs connected to your enterprise data.
LLM OCR lowered the error rate but changed what an error looks like. Why fluent output hides silent substitutions, and what has to sit around the model.
Discover why the future of OCR isn
Anthropic is an AI safety and research company that's working to build reliable, interpretable, and steerable AI systems.
Anthropic is an AI safety and research company that's working to build reliable, interpretable, and steerable AI systems.
Tips and patterns for getting the most out of Claude Code, from configuring your environment to scaling across parallel sessions.
Credentio is a newly released, open-source C++ library from Google that allows developers to integrate high-performance, local-first validation of C2PA Content Credentials into their client and server applications. By processing assets…
Deploying secure, real-time Edge AI on Raspberry Pi is now simplified using LiteRT and lightweight Gemma open models. LiteRT optimizes CPU and GPU performance, delivering fast token speeds for models like Gemma4, enabling real-time…
Conductor has evolved from a Gemini CLI extension into a portable plugin, bringing conversational Spec-Driven Development (SDD) to ecosystems like Antigravity CLI and Claude. Rather than relying on strict command sequences, developers…
Deploy NVIDIA Nemotron 3.5 ASR for low-latency, production-ready speech recognition with 6x higher throughput and multilingual support.
Cerebras CS-4 delivers up to 30x faster AI inference than GPUs, with a modular rack-scale architecture built for hyperscale AI deployment.
Cerebras powers OpenAI’s GPT-5.6 Sol Ultrafast in the OpenAI API, delivering frontier intelligence at real-time speeds for critical AI work.
Build lightning-fast multimodal AI apps with Gemma 4 on Cerebras. Learn image understanding, vision workflows, and high-speed inference for developers.
Discover why AI loops require verification to avoid spiraling. See how Cerebras runs Gemma 4 at 1,500 tokens/sec for fast, autonomous visual loops.
Reasoning boosts AI accuracy, but at a steep cost. Explore test-time compute, agent performance, speed tradeoffs, and when thinking hurts.
Kimi K2.6 on Cerebras matches Gemini 3.5 Flash in intelligence while delivering 5× faster output, lower latency, and open-weight flexibility.
Cerebras powers the world's fastest AI inference on the biggest wafer chip. Cerebras CS-4 delivers up to 30x faster inference than GPUs.
Amazon Bedrock now supports the OpenAI GPT-5.6 models, Terra and Luna, in India with India geographic cross-Region inference. If you have local data processing requirements, you can now use these models at scale while Amazon Bedrock…
OpenClaw is the fastest-growing project in GitHub history. Peter Steinberger and several maintainers share what they learned in the project's first six months. The post OpenClaw went viral. Meet the maintainers building and securing it.…
Just-in-time access, resource policies, and session auditing, right where you need them.
The Deployments page now has redesigned filters that make it faster to find a specific deployment. You can: Apply suggested filters for common searches. Type to find an option instead of scrolling through the list. Describe what you’re…
Hello you fine Internet folks,
OpenAI is expanding its presence in Brazil, deepening engagement with developers, businesses, and communities to support AI adoption across the country.
The conference with hot chips and even hotter companies
Choosing an open source LLM means weighing license terms, hardware limits, and your real workload before you ship.
A Linear project management guide covering setup, cycles, triage rules, and automating tickets into merged pull requests
The EU AI Act compliance deadline is August 2, 2026. Learn what the EU AI Act requires, and how LangSmith and LangChain OSS products help you meet each requirement.
How Factory AI uses LangSmith to debug issues and close the product feedback loop, resulting in a 2x improvement in iteration speed.
Reflections on how LangChain has evolved — including our products, ecosystem, and community — over the past two years, and where we're headed next.
Dive into LangSmith product usage patterns that show how the AI ecosystem and the way people are building LLM apps is evolving.
Harrison Chase on LangChain's 3-year journey from open source to $1.25B company, announcing LangChain 1.0, LangSmith expansion, and $125M funding.
Learn how LangChain rebuilt their chatbot using Deep Agents and subgraphs for sub-15-second responses with precise citations—and what you can apply.
Read about the latest product updates, events, and content from the LangChain team
In this post, you will learn how GoDaddy migrated from their legacy business intelligence (BI) tool to Amazon Quick. This was a two-year transformation that delivered results across every dimension of the business: 15,000 hours saved…
The SageMaker Python SDK v3 redesigns script mode with unified ModelTrainer and ModelBuilder classes. This post walks through two end-to-end examples, a scikit-learn Random Forest and a multi-GPU Stable Diffusion 3.5 LoRA fine-tune,…
We built WikiBench to test whether generated wikis help coding agents. Pairing a wiki with source code scored higher than source alone, at lower cost.
Anima Anandkumar has spent two decades in AI, from classical math to deep learning and back. Now she's using it to model the physical world, from weather to fusion reactors.
ChatGPT for Teachers is expanding to 55 U.S. school systems, bringing secure AI tools, training, and support to over 100,000 more educators and staff.
OpenAI’s new report explores how students and educators use ChatGPT to make learning more continuous, with support that extends beyond the classroom.
The fact that AI wrote 1M LOC and then refined it over the course of the next couple of months to produce a reliable piece of software that is currently running on millions of developer machines is absolutely mind blowing. And you can…
Building on a long history of HPC-focused cores, and looking beyond HPC
Muse Image from Meta Superintelligence Labs is now available on AI Gateway. It is their first image model and a separate family from Muse Spark, returning images rather than text. Send a prompt and get an image back, or send an image…
Gemini 3.5 Transcribe from Google is now available on AI Gateway for recorded and live audio: google/gemini-3.5-transcribe transcribes a complete audio file in one request. google/gemini-3.5-transcribe-live transcribes audio over a…
Discover how ClickHouse .NET Driver 1.1–1.3 adds type-safe POCO workflows, extensible serialization, broader type support, performance improvements, and official ecosystem integrations.
EVE Online: The Move to Python 3 Begins! EVE Online has been one of the most interesting case studies in Python at scale for over twenty years now. They've been running on Stackless Python since their launch in 2003, and their last…
These are the lessons we learned evaluating LLMs for real-world secret scanning. The post How to evaluate LLMs before production appeared first on The GitHub Blog .
Build a governed weekly reporting workflow with Amazon Quick Desktop and Amazon FSx for NetApp ONTAP. An Amazon S3 access point exposes an approved folder to a Quick knowledge base, and a custom skill drafts cited weekly reports and…
Speed Insights now has a free tier that gives you a high-level performance overview from your real users. The new free tier: Is available on every plan, for any number of projects Includes 10,000 events per team, every 30 days Install…
PostgreSQL 19's new `WAIT FOR` command enables read-your-writes consistency on asynchronous replicas by letting individual reads wait for a specific WAL position.
ClickGap reviews merged ClickHouse changes, executes reproducers, rejects false positives, and attributes regressions to specific commits.
Few professions are as exacting as the practice of law. A team reviewing a contract or building a case works inside strictly privileged information, firm-specific playbooks, and a body of law that changes constantly. The work thrives on…
OpenAI CFO Sarah Friar explains how advances across chips, compute, models, and products compound to deliver more useful intelligence at greater scale and lower cost.
Hello you fine Internet folks,
An industry legend starts covering the inevitable!
Hello you fine Internet folks,
OpenAI banned Russia-origin accounts using AI to promote a fake Israel-based think tank and a “sovereignty” index praising Russia and criticizing the West.
MiniMax M3 and M2.7 are free on AI Gateway via GMI Cloud through Sunday, September 6. Use minimax/minimax-m3-free or minimax/minimax-m2.7-free to route requests to GMI Cloud. These model IDs will return an error after the free period…
Wan 3.0 from Alibaba is now available on AI Gateway as alibaba/wan-v3.0-video . Wan 3.0 combines text-to-video, image-to-video, first- and last-frame conditioning, and reference-based generation in one model. References can include…
Video generation on AI Gateway can now run asynchronously. By default, generateVideo keeps one HTTP request to AI Gateway open until the result is ready. Because video generation can take seconds or minutes, that request can exceed…
Vercel Connect now includes a managed connector for Linq , so your apps and agents can send and receive messages over iMessage, RCS, and SMS. As a Vercel Managed Connector , Vercel can create a Linq account and phone number for you, or…
You can now build bots that hold end-to-end encrypted 1:1 and group conversations on XChat with the new XChat adapter for Chat SDK. The adapter handles all encryption, key management, and signature verification automatically. Bots can…
Migrate off Confluent without the “big cutover weekend.” Redpanda Shadowing carries topic data, schemas, offsets, and ACLs on a single link. Available on Self-Managed , BYOC, and Dedicated.
We built a plugin for the GitHub Accessibility Scanner to make sure your alt text is actually accessible. Here's how it works. The post Your alt text passes automated checks. That doesn’t mean it’s any good. appeared first on The GitHub…
All about making things affordable
We migrated the Cloudflare Blog to EmDash to prove our stack at massive scale. Here is how we stress-tested performance, safely routed production traffic, and redesigned the frontend experience.
Release: llm-anthropic 0.27 This release of the Anthropic plugin for LLM mainly provides compatibility with the recently released anthropic v1.0.0 Python library, which switches from httpx to httpx2 . OpenAI made the same change in…
The massive growth of generative AI has fundamentally altered data center design. As distributed model training scales to span hundreds of thousands of GPUs,...
AI factories are power-constrained industrial systems. The question is no longer how many GPUs fit in a data center, but how much AI output each available...
A preview of what we’re building—and what we’ll share at TailscaleUp.
Banking has become an always-on business. Customers expect instant payments, accurate balances, and continuous access to financial services...
And this was a debug session from hell, enormously helped by an AI doing much of the grunt-work. I'd like to call it my tireless helper, but the AI several times stated flat out that this was impossible and unsolvable and that we should…
Release: llm 0.33 My highlights from this release: Upgraded to the OpenAI Python library 3.x and switched the HTTP client dependency from httpx to httpx2 . #1608 , #1631 I shipped a quick 0.32.1 fix for this yesterday, but this is the…
Did you think RSI stopped at model training?
Simile’s CEO about his journey from the viral Generative Agents to creating 8 Billion Digital Twins of every living human... and why it’s gone from fun exploration to very serious business.
Cloudflare's new Bot Preference Sync automatically aligns your robots.txt file with your AI bot policies for Search, Agent, and Training. Easily manage which bots access your content without maintaining static files.
Release: llm 0.32.1 Fresh installs of LLM stopped working the other day because the OpenAI Python library dropped its usage of httpx , and it turned out LLM depended on that library but only installed it via a transitive openai…
Release: llm-openrouter 0.7 Now that this plugin is compatible with LLM 0.32 it can display the reasoning traces for LLMs available through OpenRouter. Updated for compatibility with LLM 0.32 . Models now use OpenRouter's implementation…
Stop Making TUIs Thomas Ptacek advocates for building real native user interfaces for even the smallest of personal tools, because coding agents have reduced the cost of getting a usable-enough GUI up and running to almost nothing. I…
Samuel Yeboah , Francesco Di Chiara and Mingliang Liu Today, Netflix runs two Flink autoscalers. That is exactly one more than we want. We built the first one in-house years ago, when there was no mature option suited to our platform.…
After I released version 1.0, I figured I would have to do the rotations myself. So I sat down with ChatGPT and I didn’t get it to write the code, but I got it to educate me. With a patient, interactive tutor, I was able to finally do…
Welcome to the August 2026 ClickHouse newsletter, which will round up what’s happened in real-time data warehouses over the last month.
Yes, we’re confused too.
ChatGPT search now uses the site:operator at scale Promptwatch is part of the emerging "GEO" space, for Generative Engine Optimization - the chatbot version of SEO, where companies offer tools and consulting to help your site increase…
Since announcing Google Antigravity in Gemini Enterprise Agent Platform at I/O in May, we’ve heard helpful feedback from our customers. Your developers want easy access to coding agents across surfaces. Your enterprise governance team…
Cloudflare OAuth now supports optional scopes, giving users more control over what an app can access and helping developers build secure consent flows around the task at hand.
Recommender systems (RecSys) are one of the most ubiquitous machine learning problems in the consumer internet industry yet notoriously difficult to train and...
It’s never been easier to start an AI-powered startup on Google Cloud. You grab an API key from Google AI Studio at breakfast, paste it into Antigravity, and by lunch you’ll have a nascent prototype of your product. But it’s not all one…
Introducing Intelligence Age, a new OpenAI blog exploring how transformative AI could reshape power, governance, the economy, and individual freedom.
Introducing Intelligence Age, a new OpenAI blog exploring how transformative AI could reshape power, governance, the economy, and individual freedom.
Last week I recorded an episode of the Talking Postgres podcast with Claire Giordano on the subject of "How AI is changing software development". We had a really great conversation. Here are a couple of my highlights from a lightly…
OpenAI reaffirms Zero Data Retention for eligible API customers and previews Private Safety Processing for advanced AI safety without compromising data privacy.
Modern vision-language models (VLMs) can support tasks such as visual question answering, captioning, and image-text reasoning. In practice, however, the data...
Robots need policies that can adapt to their sensors, environments, and tasks while running on onboard computing hardware. World models offer a foundation for...
Replit introduces Free Mode, powered by GPT-5.6 Luna, so anyone can turn ideas into working software without worrying about token costs.
ChatGPT Ads is expanding to 31 European markets. Learn how advertisers can reach people as they explore, compare options, and make decisions.
Every enterprise has valuable data trapped in messy, unstructured documents. Today,...
OpenAI is strengthening monitoring, alignment, and security for frontier AI models. See how new safeguards are guiding the pace of model development.
ChatGPT for Teens helps teens learn, think critically, and use AI with confidence, with stronger built-in protections, healthy-use features, and additional controls for parents.
NVIDIA teams use ChatGPT Work to reduce manual tasks, connect fast-moving signals, and scale successful workflows globally.
No GPUs, no Agents, just really, really, really good infra and distribution.
Machine learning models are only as good as the signals they receive. A fraud detection...
Nvidia wants you building your own model, not buying from Anthropic/OpenAI.
AI is reshaping cybersecurity for attackers and defenders alike. Learn how OpenAI is strengthening its defenses and what security teams can do now.
OpenAI joins PORTS-Pike project, expanding community investment and supporting thousands of Southern Ohio jobs
Shadow traffic proves a candidate is operationally sound. It can't tell you if users like it better. Run the split at the endpoint instead of in your app code.
I started building my markdown-svg-renderer tool in May , but I've since added enough features to it that it's worth talking about here again. It's evolved into my ideal tool for sharing Markdown transcripts that include SVG documents.…
Friday's big release was Qwen 3.8 27B , an Apache 2 licensed 27B parameter vision-capable LLM from Alibaba's Qwen research lab. I've been looking forward to this one: 27B is an excellent size for running a model on a reasonably specced…
Hint: It’s really not a distillation story.
Down, but not out!
Cloudflare's data shows a clear impact on Internet traffic from Iceland to Spain and Portugal, following the path of totality of the total solar eclipse that occurred on August 12, 2026.
The GitHub Universe session catalog is live. Explore interactive workshops, community talks, demos, and panels. Plus, register before August 19 to save $300. The post Your guide to GitHub Universe 2026 is here: The schedule just…
AI teammate category just had its most significant new entrant yet
Enterprise content is no longer just something people consume. As organizations increasingly rely on AI to extract and act on information from documents, images, audio, and video, Azure Content Understanding is expanding support for the…
Alibaba released the open weights for Qwen3.8-2.4T-A95B (Qwen3.8-Max), its largest open-weight model, bringing near-frontier capabilities to the open...
SQLite corruption? Sorted.
Reflections on AI's writing ability and how AI models get more capable.
WhatsApp is committed to helping people stay safe while protecting the privacy of their messages. As scam tactics evolve — from impersonation to social engineering to AI-generated lures — we’re always evolving as well, so that our…
Pharma is suddenly paying for Bio × AI tools, and Chai is leading the pack with four deals closed this summer. Cofounder Matt McPartlon and Product leader Neil Patil explain why.
Discover what's new in pg_clickhouse v0.10.0: 16 of 22 TPC-H queries now push down in full, plus a rebuilt C driver and broader aggregate support.
Hello you fine Internet folks,
Inside Medium’s move from relational features to list features in its ScyllaDB-based feature store “Keep readers reading” is the not-so-simple goal of Medium’s recommendations system. To predict what’s most likely to appeal to a…
Explore ClickStack’s latest upgrades, from richer trace navigation and Prometheus connectivity to smarter dashboards, quieter alerts, faster filtering and support for exponential histogram metrics.
Meet Susie Solis, a Technical Support Engineer on the Customer Experience team here at ScyllaDB.
Musings on model alignment, what determines safety, and where we go from here.
Cloudflare Radar Researcher is a new AI-powered tool that lets you explore global Internet trends and traffic data using plain language. Built entirely on Cloudflare's Developer Platform, it turns natural language queries into real,…
We are launching updated community programs, including Cloudflare Ambassadors and Community Engineers, backed by $1M in open-source funding. Learn how we are supporting maintainers and scaling our developer community.
Keep AI agents useful without combining their riskiest capabilities.
This release brings speedups for GROUP BY ... ORDER BY ... LIMIT, three JOIN improvements, four vector search improvements, position-aware phrase search, EXPLAIN ANALYZE, unified URL access, and more!
Physical Intelligence runs both its OLAP and OLTP workloads on ClickHouse managed Postgres and ClickHouse Cloud
Hello you fine Internet folks,
How ClickHouse Managed Postgres uses WAL-aware backpressure to slow client writes, protect disk space, and let the archiver recover.
Fireworks is the first and only dedicated inference platform Voyage AI by MongoDB has partnered with. The full Voyage lineup now runs natively on Fireworks: the Voyage 4 family, voyage-multimodal-3.5, and rerank-2.5.
Martin Kleppmann and Chris Riccomini's scalability considerations for designing data-intensive applications -- from the second edition of the Designing Data-Intensive Applications book
Scaling our curation and measurement of the open ecosystem.
Arm’s 5-series cores are meant for tasks where performance barely matters, but power and area efficiency are top priorities.
Capacity to train strong models is proliferating.
Kimi K3 is the first open 3T-class model. See how it benchmarks, what it costs, and how to call it on the Together AI API, with copy-paste code examples.
Tailscale wasn’t exploited. We still should have stopped the intrusion.
by Aarti Laddha , Richard Diaz-Cool , Rishika Idnani , Venkatesh Selveraj Netflix supports a vast and evolving set of features and content types, ranging from 4K streaming and immersive audio to live streaming and cloud gaming, across a…
Hello you fine Internet folks,
Authors: Ying Li , Arjun Rao , Shradha Sehgal Introduction Recommendations sit at the heart of the Netflix experience. Our current production models rely on thousands of hand‑crafted features over users, items, and interactions, along…
Oxide joins Anthropic's Project Glasswing, applying Claude Mythos 5 to find and patch vulnerabilities across its open-source stack, firmware to network.
Informational only, not legal advice. Confirm all regulatory details against the...
Hello you fine Internet folks,
Kimi K3, a 2.8 trillion parameter multimodal model by Moonshot, along with a custom-trained DFlash speculator, is now available on Modal.
Trilogy’s new AI Cybersecurity Playbook centers on Kimi K3, leveraging open weights to ensure defenders have reliable, high-volume inference options. By moving beyond closed models, organizations gain the deployment flexibility—from…
Kimi K3 and Thinking Machines' Inkling shipped open weights this week. What self-hosting a frontier-class coding model buys you on your own infrastructure, and what it costs.
Kimi K3 and Thinking Machines' Inkling shipped open weights this week. What self-hosting a frontier-class coding model buys you on your own infrastructure, and what it costs.
Editor’s Note (7/25/2026): The article has been edited with more information about the L2 behavior along with the bandwidth of the die to die interface.
Today, we’re welcoming The Interaction Company of California, the makers of Poke, to Cognition.
A first look at the new visual Pipeline Builder for Redpanda Connect (preview), plus a roundup of recent CDC and connector updates.
How to prioritize the things that matter most for planning, executing and de-risking your NoSQL database migration
Across Databricks, thousands of customers build production workloads that map freeform...
The global implications on the AI ecosystem.
Riviera is the Dropbox content processing platform that’s been iteratively improving content transformation in our products for roughly a decade.
Hello you fine Internet folks, this article is a sequel to the Scrying the AMD GFX1250 LLVM Tea Leaves article where we are going to look at the differences between GFX1250 and GFX1251.
Tracking data is now the richest signal in sport, but the real gap is turning the...
Take Home Assistant beyond your home network with Tailscale.
You're standing at the register. You tap your card. A tiny spinner appears for maybe...
IntroductionApache Spark 4.2 moves more of the modern data and AI stack into the...
How (and why) we built a scheduling system that can scale to 1 million concurrent sandboxes (per workspace) in seconds.
Hierarchical Interest Representation is a research area for Meta Ads. We’re exploring an upstream representation layer over the universe of Ads entities – users, advertisers, products, services – learning unified embeddings that connect…
Inkling, a general-purpose multimodal model by Thinking Machines, along with a custom trained DFlash speculator, is now available on Modal.
Use state-of-the-art, open-source LLMs and image models at blazing fast speed, or fine-tune and deploy your own at no additional cost with Fireworks AI!
BetterTracker eliminated a standalone vector database by running OLTP and vector search together on CockroachDB—cutting costs, complexity, and compliance risk in one move.
One year ago, Cognition and Windsurf came together over the craziest 72 hours of our lives. Here’s what we’ve built together since.
CPU designs must evolve to keep up with changing workloads.
Full CDC semantics in Iceberg mean the lakehouse reflects what the source database looks like right now, not what it looked like during last night’s batch window. Get to know the latest addition to Redpanda Connect.
This post establishes a reusable pattern for operational workloads that genuinely move the needle: fraud detection...
Fable 5 costs twice as much per token as Opus 4.8. But when we ran both models on FrontierCode 1.1 using our new Fusion architecture, Fable cost less —…
AI agents aren't just changing how applications interact with databases, they're changing how teams operate them.
The most serious test to date of open source AI’s viability is happening right now.
What is Modal? A machine that runs programs of arithmetic & logic operations on information.
Azure Synapse has served as a reliable foundation for SQL analytics at scale, and...
Modern AI applications are no longer single-shot inference calls. They are long-running agents that plan, act, observe, and retry across time.
Claude is now generally available in Microsoft Foundry. Here's everything else that shipped between Build 2026 and the end of June — autopilot agents, expanded Toolboxes and Routines, Agent Optimizer's private preview, and more. The…
A hands-on walkthrough of the new Oracle input in Redpanda Connect. No Debezium, no Kafka Connect runtime, no JVM.
Wide-column flexibility doesn’t have to come at the expense of performance -- see where the two models differ, where each one wins, and why you no longer have to choose
A step-by-step tutorial on how to stream every insert, update, and delete from your database using MySQL and a faster, simpler alternative to Kafka Connect.
One tiny module, two ways to win on Diffusion LMs
We raised $800M to accelerate the shift to open-source AI. Here's why the economics of closed models don't scale, and what we're building next.
This year marks Meta’s 10th consecutive year as a sponsor of the Python Software Foundation (PSF), the charitable organization dedicated to advancing, supporting, and protecting the open-source Python programming language and the…
Announcing our integration with Claude Science, bringing Modal's elastic compute to researchers when they need it.
Nine papers at ICML 2026 across the full stack. The research that becomes the Together platform. Find us at booth B714 in Seoul.
Use state-of-the-art, open-source LLMs and image models at blazing fast speed, or fine-tune and deploy your own at no additional cost with Fireworks AI!
Authors: Lequn Wang , J iangwei Pan , and Linas Baltrunas Figure 1. Autoregressive homepage generation. GenPage builds a Netflix homepage one row or entity at a time, each one conditioned on what’s already on the page and the user’s…
An assessment of the open ecosystem and the motivations behind releasing models
A utility company deploys drones to inspect hundreds of miles of power lines. A police...
Hello you fine Internet folks,
There's a problem with Apache Kafka's log compaction. Here's what we found, how to reproduce it, and how we solved it in Redpanda.
By Zhuoning Yuan , Ta-Ying Cheng , Benjamin Klein , Bahareh Azarnoush Introduction At Netflix, we build technology to help storytellers bring their creative visions to life and to help members discover the stories they love. To connect…
This post was originally an op-ed co-authored with Kevin Xu of Interconnected for a general, non-technical audience.
Tuning network performance across the layers of a stack you build end to end
About 3 years since I started writing weekly.
Next-generation artificial intelligence (AI) and high-performance computing (HPC) chips routinely exceed 1000 W Thermal Design Power (TDP). Simply put, standard air cooling cannot manage these extreme heat loads. The alternative —…
Avoid AI lock-in with Aperture’s flexible, identity-aware AI stack
It's a one-way door and we weren't ready for it.
Tailscale on jailbroken Kindles now supports proxies and SSH. And a plugin supports Kobos and Pocket Readers, too.
Hello you fine Internet folks, today we have an interview with Kira Boyko, the Product Director of Intel Xeon 6+.
Meet Faisal Saeed, Principal Customer Engineer on the Customer Experience team here at ScyllaDB.
Today’s coding benchmarks have established that models can write correct code, but the question we should really be asking is: can models actually write…
Why edge AI development is still hard AI is no longer confined to cloud experiments. Developers are increasingly expected to deliver AI inside apps, devices, and edge systems where responsiveness, privacy, resilience, and local control…
Learn how new Document Translation capabilities in Azure Translator, available in Foundry Tools, help developers translate images, PDFs, Office files, DITA, XLIFF, and future LLM-powered document workflows. The post Expanding the Reach…
Microsoft Foundry Managed Compute is a new GPU platform-as-a-service for hosting open-source and custom AI models behind the same endpoint, SDKs, and bill as frontier models. The post Announcing Foundry Managed Compute: Run open models…
How our Rust collections library defends against adversarial trait implementations.
This was my last week at the Allen Institute for AI (Ai2), where I got the great privilege to work on the Olmo models, to grow, to learn, and to have broad lasting impacts.
Where marginally higher intelligence drives value, and where it doesn't.
Cognition has raised over $1B at a $26B valuation, led by Lux Capital, General Catalyst, and 8VC.
Over the last 25 years of building Google’s global network, we’ve navigated major architectural eras — from the Internet, to streaming, and the cloud. Today, we are squarely in the midst of a fourth: the AI era. The applications in the…
Gemini Flash 3.5, Mythos, open-closed balance, America's open-source surge, emerging power struggles and more.
We've raised $355M at a $4.65B valuation to continue building the production cloud for AI.
When most people think of A/B experimentation, they think of button colors, landing page layouts, or checkout flows. At Google, many fundamental infrastructure improvements also need the rigor of A/B experimentation. Optimizing a memory…
llm-d v0.7 shifts focus from proving capabilities to making them deployable, with changes across deployment tooling, hardware support, documentation, and continuous integration.
The index covers 45 sources from the watchlist. 37 publish a feed and are read from it; the other 8 publish none, so their index pages are scraped and each new post's own page supplies the title and date its card omits.
Last run finished 28 Aug 2026 at 00:08 UTC. The index holds 710 posts, keeps them for 120 days, and shows 236 posts on this page.
2 sources failed on the last run. Netflix TechBlog (HTTP 429); Replit (HTTP 403)
The last run read 833 posts across every source and added nothing new.
4 sources on the watchlist are not aggregated here, so check them by hand.
A date shown as "first seen" is not a publication date. Some sources publish no date at all, so the page records when the post entered the index instead of guessing.