Straight from the source · No aggregators, newsletters or social media

The Tech Gazette

Vol. I No. 2 ★★★ Wednesday, September 23, 2026 ★★★ Closing 8:47 AM BRT

News

Announcements from the labs, companies and platforms themselves.

Wednesday, September 30

  1. VS Code2:00 PMcode.visualstudio.com
    Visual Studio Code 1.140 (Insiders)

    Learn what's new in Visual Studio Code 1.140 (Insiders) Read the full article

Today

  1. VS Code2:00 PMcode.visualstudio.com
    Visual Studio Code 1.139

    Learn what is new in Visual Studio Code 1.139 Read the full article

Yesterday

  1. LangChain11:32 PMlangchain.com
    What Is Jev? A Guide to TypeSafe AI’s System One Model

    What is Jev? Learn how TypeSafe AI’s System One model makes fast, structured decisions, where it fits in the agent loop, and how to use Jev with LangChain

  2. NVIDIA11:30 PMblogs.nvidia.com
    At AI Day Singapore, NVIDIA and Partners Showcase AI Advancements Across Southeast Asia

    NVIDIA AI Day Singapore, which takes place Sept. 22-23 at the Raffles City Convention Centre, is offering attendees opportunities to explore the hands-on training, expert-led sessions and advanced tools to accelerate their work in AI and high-performance computing. At the event,…

  3. GitHub Changelog11:14 PMgithub.blog
    OpenTelemetry in the GitHub Copilot app

    Understand how Copilot agents perform and interact with models and tools. The GitHub Copilot app now supports OpenTelemetry (OTel) configuration through enterprise-managed settings. OTel is an open source observability framework.… The post OpenTelemetry in the GitHub Copilo…

  4. GitHub Changelog9:34 PMgithub.blog
    New features and improvements in Copilot for JetBrains

    GitHub Copilot for JetBrains 1.18.0 brings AI-assisted tool approvals, more control over agent conversations, and shared skills and instructions for your organization. You can also review plans with the Codex… The post New features and improvements in Copilot for JetBrains…

  5. OpenAI9:00 PMopenai.com
    Grab and OpenAI bring practical AI skills to Southeast Asia

    OpenAI and Grab launch GO Forward with AI, a regional programme helping 30,000 partners build practical AI skills across Southeast Asia.

  6. GitHub Changelog7:24 PMgithub.blog
    Faster C++ code intelligence with whole codebase indexing

    C++ code intelligence in GitHub Copilot CLI is now faster with support for whole codebase indexing. C++ repositories can contain millions of lines of code across deeply connected source files… The post Faster C++ code intelligence with whole codebase indexing appeared first…

  7. AWS News7:23 PMaws.amazon.com
    Introducing Amazon CloudWatch Omni: collaborative AI-powered observability for your applications

    Amazon CloudWatch Omni is the next evolution of CloudWatch — unified observability that brings your applications and AI agents into one reimagined experience, with auto-discovered topology, natural language queries, and AI-guided investigation powered by AWS DevOps Agent.

  8. AWS News7:21 PMaws.amazon.com
    Introducing Amazon CloudWatch Omni: AI-powered observability for generative AI and agentic workloads

    Learn how Amazon CloudWatch Omni delivers AI-powered observability purpose-built for generative AI and agentic workloads. Trace, evaluate, and experiment with AI agents across any framework—directly from your IDE or a standalone web experience—using open standards and built-in ev…

  9. OpenAI6:00 PMopenai.com
    Better prompt caching for GPT-6

    Learn how GPT-6 improves prompt caching with higher cache hit rates, new diagnostics, explicit breakpoints, and controls that reduce latency and costs.

  10. OpenAI3:00 PMopenai.com
    Introducing GPT-6 Sol and Luna

    Meet GPT-6 Sol and Luna, two models that bring frontier intelligence to everyday work with different balances of capability and cost.

  11. Kubernetes3:00 PMkubernetes.io
    Spotlight on SIG Apps

    As Kubernetes adoption has grown, the conversation has shifted beyond running containers to managing increasingly complex application lifecycles. Modern platforms support stateless web services, stateful databases, batch processing, AI workloads, and platform services. At the sam…

  12. Cohere2:45 PMcohere.com
    AI Change Management

    Learn what AI change management involves and how enterprises can adapt roles, workflows, and governance as adoption scales.

  13. GitHub Changelog2:10 PMgithub.blog
    Claude Opus 5.5 is now available in GitHub Copilot

    Claude Opus 5.5, Anthropic’s newest Opus model, is now available in GitHub Copilot. You can use it for agentic coding, long-running agentic tasks, and knowledge work. In early testing, Opus… The post Claude Opus 5.5 is now available in GitHub Copilot appeared first on…

  14. LangChain2:08 PMlangchain.com
    The Reliability Layer for Healthcare AI: Common LangSmith Use Cases

    See how LangSmith helps healthcare AI teams turn clinical review into reusable evaluators, datasets, and release gates for safer AI in production.

  15. GitHub Changelog2:00 PMgithub.blog
    OpenAI’s GPT-6 Sol and GPT-6 Luna now available

    OpenAI’s GPT-6 family is expanding in GitHub Copilot with two additional models: GPT-6 Sol, and GPT-6 Luna. Joining the previously released GPT-6 Astra, these new options let you select the… The post OpenAI’s GPT-6 Sol and GPT-6 Luna now available appeared first…

  16. Databricks1:00 PMdatabricks.com
    Genie One MCP: Give any AI Agent the Right Business Context

    Business leaders often have access to plenty of data, but still can’t get a reliable...

  17. Databricks1:00 PMdatabricks.com
    The Genie One MCP is now Generally Available

    AI coworkers and coding agents are spreading fast across organizations, and each...

  18. Docker12:06 PMdocker.com
    Meet the Ecosystem: Partners and Customers at WeAreDevelopers with Docker

    Meet the partners and customers bringing practical AI, security, and development sessions to the Docker Pavilion at WeAreDevelopers. The post explains why a strong ecosystem matters to developers, announces the sessions and speakers, and invites attendees to connect with the team…

  19. GitHub Changelog11:11 AMgithub.blog
    Security improvements for SSH

    We’re removing several SSH algorithms, adding a new algorithm, and requiring larger RSA SSH keys to improve security. The changes are as follows: We’re removing the ability to use RSA… The post Security improvements for SSH appeared first on The GitHub Blog.

  20. Cloudflare11:04 AMblog.cloudflare.com
    We just shipped support for the ugliest part of HTTP: Vary

    Vary support is now available in Cache Rules on every plan. You can normalize known negotiation headers, pass exact values through to the origin when those small differences matter, or bypass cache when the variation is too unpredictable.

  21. JetBrains10:16 AMblog.jetbrains.com
    Code Quality Q&A With the JetBrains Qodana Team

    Code quality is shaped by countless decisions, from how developers manage complexity to how quickly they detect regressions. But common concepts such as technical debt, cognitive complexity and unit testing are not always clearly understood. In a new code quality Q&A session,…

  22. Cloudflare10:00 AMblog.cloudflare.com
    Introducing Worker Previews: Isolated preview environments for every change your agent makes

    Worker Previews gives every branch its own URL, configuration, state, and observability, so you and your agents can test changes in parallel without affecting production.

  23. NVIDIA9:00 AMblogs.nvidia.com
    NVIDIA Isaac ROS 5.0 Advances Agentic, Open Source Robotics Development

    To build and deploy sophisticated robotics applications that can perceive, reason and act in dynamic environments, developers need new physical AI models and tools. The ROS open framework is a project from Open Robotics that helps humans build robots. NVIDIA Isaac ROS 5.0 — a col…

  24. OpenAI9:00 AMopenai.com
    Parallel cut research time and cost in half with GPT‑6 Astra

    GPT‑6 Astra allowed Parallel’s agents to research and synthesize labor-market data in half the time and at half the cost vs. prior models.

  25. GitHub Changelog6:21 AMgithub.blog
    Deprecation notice: All-platform CodeQL bundle

    Starting with CodeQL CLI 2.27.0, the all-platform CodeQL bundle (i.e., codeql-bundle.tar.gz and codeql-bundle.tar.zst), which includes the binaries for all supported platforms up to this release, is marked as deprecated. In… The post Deprecation notice: All-platform CodeQL…

  26. Node.js6:03 AMnodejs.org
    Node.js 26.10.0 (Current)
  27. JetBrains6:00 AMblog.jetbrains.com
    JetBrains Air: Building a System of Products for Agentic Software Development

    AI can produce code. Organizations still have to produce software. Agentic development is changing how software gets made, but it hasn’t changed what it costs to be wrong. Six months ago, we began publicly experimenting with agentic development environments. Around the same time,…

Monday, September 21

  1. GitHub Changelog10:25 PMgithub.blog
    Refreshed repository pull requests page generally available

    The new repository pull requests page is now generally available to all GitHub users. Highlights The new page makes it easier to find and act on pull requests in a… The post Refreshed repository pull requests page generally available appeared first on The GitHub Blog.

  2. OpenAI9:00 PMopenai.com
    Priorities and principles for effective third party assessments

    OpenAI outlines priorities and principles for rigorous, secure, and independent third-party AI safety assessments of frontier models and safeguards.

  3. Hugging Face9:00 PMhuggingface.co
    How UK AISI and EvalEval Are Making Benchmark Results Reproducible
  4. Hugging Face9:00 PMhuggingface.co
    Transformers now runs llama.cpp quants
  5. Hugging Face9:00 PMhuggingface.co
    Jun Kim, oMLX creator and maintainer, joins Hugging Face to support the MLX community
  6. Together AI9:00 PMtogether.ai
    Canary rollouts: upgrade models in production without downtime

    A hard model swap exposes every user at once, and rolling back means cold-starting the old deployment under pressure. Here's how staged traffic ramps, metric gates, and automatic rollback work on dedicated inference.

  7. xAI9:00 PMx.ai
    How SpaceXAI is using Grok Bot to scale customer support

    We rebuilt the combined SpaceXAI and Cursor support operation around Grok Bot, expanding to a much broader product portfolio without adding headcount.

  8. Fireworks AI9:00 PMfireworks.ai
    Introducing The Specialized Intelligence Index

    Fireworks introduces the Specialized Intelligence Index, a one-stop destination for real-work benchmarks across industries.

  9. Vercel9:00 PMvercel.com
    GPT-6 Sol and Luna now available on AI Gateway

    GPT-6 Sol and GPT-6 Luna from OpenAI are now available on AI Gateway. Both models bring GPT-6 improvements in professional work, coding, computer use, factuality, and communication at a lower price than GPT-6 Astra. GPT-6 Sol ( openai/gpt-6-sol ) is suited to complex professional…

  10. Vercel9:00 PMvercel.com
    Drives for Vercel Sandbox are now in public beta

    Drives for Vercel Sandbox are now available in public beta on Hobby, Pro, and Enterprise. A Drive is persistent storage that you mount as a directory in a Vercel Sandbox. It isn’t tied to a single sandbox, so you can reuse the same Drive across runs and different sandbox instance…

  11. Vercel9:00 PMvercel.com
    Claude Opus 5.5 now available on AI Gateway

    Claude Opus 5.5 from Anthropic is now available on AI Gateway. It is a step-change improvement over Opus 5, with its biggest gains in agentic coding, long-running agent tasks, and knowledge work. Anthropic cites that Opus 5.5 performs at the level of Fable 5.1, but ~30% faster an…

  12. Netlify9:00 PMnetlify.com
    Security Update: Critical Next.js vulnerability in ImageResponse

    The Next.js team has disclosed a critical severity vulnerability in an upstream dependency that can lead to remote code execution when ImageResponse renders untrusted input. It is patched in 15.5.26 and 16.3.6. Applications that do not pass untrusted input into ImageResponse are…

  13. Netlify9:00 PMnetlify.com
    Claude Opus 5.5 now available in AI Gateway and Agent Runners

    Anthropic’s Claude Opus 5.5 model is now available through Netlify’s AI Gateway and Agent Runners with zero configuration required. Use the Anthropic SDK directly in your Netlify Functions without managing API keys or authentication. AI Gateway handles everything automatically. H…

  14. Netlify9:00 PMnetlify.com
    GPT-6 Sol and GPT-6 Luna now available in AI Gateway and Agent Runners

    OpenAI’s GPT-6 Sol and GPT-6 Luna models are now available through Netlify’s AI Gateway and Agent Runners with zero configuration required. Use the OpenAI SDK directly in your Netlify Functions without managing API keys or authentication. AI Gateway handles everything automatical…

  15. PostgreSQL9:00 PMpostgresql.org
    SQL Manager for PostgreSQL 7.0: meet the AI Assistant

    This is a major release, and the headline feature is our new AI Assistant. We're starting to roll out AI capabilities across our database tools, and SQL Manager for PostgreSQL is the first to get them. It's also available in SQL Management Studio for PostgreSQL, which ships with…

  16. PostgreSQL9:00 PMpostgresql.org
    pgsql-test: Real Postgres Testing for Faster Development Loops

    Application code has fast testing loops: runners, fixtures, and red-green feedback in JavaScript, TypeScript, and Python. Postgres can be tested too, but database logic often sits outside those loops. Developers have to provision state, manage transactions, or fall back to a past…

  17. PostgreSQL9:00 PMpostgresql.org
    PostgresCompare 2.2.0 Released

    PostgresCompare is pleased to announce the release of version 2.2.0, the first public release since 1.2.2 and the largest change in the product's history. The application has been rebuilt on a new foundation, and the 2.x series adds database pipelines, ad-hoc comparison, data-cha…

  18. Stripe9:00 PMstripe.com
    New trends in global card fraud: How 3D Secure and regional mandates are affecting risk

    We analyzed billions of transactions on Stripe from January 2022 to March 2026 to understand how card fraud patterns differ by region and country, what's driving those differences, and how businesses can respond.

  19. GitLab9:00 PMabout.gitlab.com
    How to design GitLab for enterprise scale

    At enterprise scale, even small architecture choices can have outsized consequences. A deployment that works for a handful of teams can become a constraint once thousands of developers, repositories, and pipelines depend on it. That makes each decision made before rollout especia…

  20. GitLab9:00 PMabout.gitlab.com
    How GitLab reduced code-per-agentic-flow ratio by 45%

    GitLab Duo Agent Platform orchestrates and automates complex tasks through agentic flows. A key part of the platform is the Flow Registry, a declarative configuration framework, built from reusable components, that compiles YAML into fully functional LangGraph flows. By using Flo…

  21. Rust9:00 PMblog.rust-lang.org
    Announcing a Maintainer in Residence: Scott Schafer for the Cargo team

    At the end of August, we announced our first Maintainers in Residence, Rust Project contributors who are funded for their upstream contributions and maintenance work from the Rust Foundation Maintainers Fund (RFMF). Since then, the Rust Leadership Council has dedicated more funds…

  22. GitHub Changelog6:13 PMgithub.blog
    GitHub Enterprise adds credential inventory exports

    Enterprise owners can now export a complete inventory of every credential that can access their enterprise (e.g., SSH keys, classic and fine-grained personal access tokens, OAuth App access tokens, and… The post GitHub Enterprise adds credential inventory exports appeared f…

  23. Amazon Science4:31 PMamazon.science
    Amazon launches research initiative with Stanford University to advance AI and science

    The collaboration aims to advance research while broadening participation and translating discovery into real-world solutions.

  24. Kubernetes3:30 PMkubernetes.io
    Kubernetes v1.37: Tracking When a PersistentVolumeClaim Was Last Used (Beta)

    Kubernetes v1.37 promotes the PersistentVolumeClaimUnusedSinceTime feature gate to Beta (enabled by default). With this feature, the PersistentVolumeClaim (PVC) protection controller adds an Unused condition to each PVC, telling you whether any running pod currently references it…

  25. NVIDIA3:00 PMblogs.nvidia.com
    NVIDIA Launches DSX Ready to Qualify Power and Cooling Products for AI Factories

    Every AI factory needs power and cooling that fit its computing architecture. As AI infrastructure expands, power, cooling, water, site and grid constraints are shaping what builders can deploy. Choosing products that fit the complete factory design helps builders turn computing…

  26. Amazon Science2:51 PMamazon.science
    Advancing AI for biology: Teaching models to design and characterize antibodies

    Three new papers from Amazon Bio Discovery address bottlenecks in AI-driven antibody engineering, from benchmarking binding predictors to experimentally validating de novo design.

  27. LangChain2:14 PMlangchain.com
    Jev is now available in LangSmith Evals

    Use Jev as a judge for LangSmith evals to evaluate agent traces with faster, cheaper structured feedback across production runs, datasets, and regression tests.

  28. Vercel2:10 PMvercel.com
    Vercel Connect now supports Microsoft Teams

    Vercel Connect now includes a managed connector for Microsoft Teams. Creating one gives your organization a Teams bot that your apps and agents run. People can @mention it in channels or message it directly, and your code receives the message and replies as the bot. As a Vercel M…

  29. AWS News1:11 PMaws.amazon.com
    AWS Weekly Roundup: AWS Builder Center mobile apps, Amazon Connect Talent GA, Amazon Corretto 27, and more (September 21, 2026)

    Living in the Netherlands, I spend a fair amount of time on trains, and that is usually where I catch up on what the builder community is writing. Until now, that meant opening a laptop or squinting at a browser tab on my phone. This week I found myself scrolling through trending…

  30. NVIDIA1:00 PMblogs.nvidia.com
    Why Deploying Physical AI at Scale Demands Safety at Every Layer

    Physical AI is moving rapidly from research to large-scale deployment. By 2035, ABI Research projects an installed base of 49 million level 3-5 autonomous vehicles (AVs), while Omdia estimates that roughly 60 million industrial robots will be deployed between 2026 and 2035. As th…

  31. Meta Engineering1:00 PMengineering.fb.com
    Open-Sourcing Rebalancer: A Generic, High-Performance Library for Solving Assignment Problems

    We’re open-sourcing Rebalancer, the assignment-problem solver that has been used to solve resource allocation problems throughout Meta for over nine years. Rebalancer separates several related concerns: how to specify an assignment problem, how to store it efficiently in me…

  32. NVIDIA1:00 PMblogs.nvidia.com
    From Enablement to Execution, Egypt’s AI Ecosystem Reaches Production Scale

    Today, Egypt’s AI builders gathered in the Grand Egyptian Museum for a reception that highlighted the nation’s rapidly growing AI ecosystem — spanning AI natives, developers, researchers, startups and enterprises — building applications across industries. The event included a key…

  33. Google Cloud1:00 PMcloud.google.com
    Global AI routing with <1% overhead on multi-cluster GKE Inference Gateway

    Demand for AI infrastructure is at an all-time high. Global accelerator shortages mean engineering teams can rarely get all the compute they need from just one data center — capacity comes a cluster here, a cluster there, often an ocean apart. At the same time, workloads are gett…

  34. Google Cloud1:00 PMcloud.google.com
    Maximizing Apache Spark availability: Mitigating compute stockouts with flexible VMs and other best practices

    The surge in AI development has created unprecedented demand for compute capacity around the globe. This can have negative implications for data processing and pipelines with Apache Spark. Whether you are managing your own Spark infrastructure or using a managed service, you can…

  35. Google Cloud1:00 PMcloud.google.com
    Scale your AI workloads faster and more efficiently with GKE Pod snapshots

    When running modern AI workloads, there’s often a conflict between performance and cost. Workloads like large language models (LLMs) load massive files, and may serve thousands of AI agents that need to execute code instantly. If each component is starting “cold” with a full data…

  36. Google Cloud1:00 PMcloud.google.com
    Strengthen your CI/CD pipeline with new Secure Source Manager capabilities

    A resilient software supply chain is the foundation of modern delivery, and securing your continuous integration and continuous delivery (CI/CD) pipeline is what keeps innovation moving safely. Notable supply chain attacks more than doubled in the first half of 2026 compared to t…

  37. Microsoft Research12:30 PMmicrosoft.com
    Improving synthesis prediction of small molecules at scale with RetroChimera

    Custom-made molecules are advancing medicine, materials, and agriculture, but producing them is slow and expensive. A new Nature paper highlights RetroChimera, a predictive model that helps accelerate chemical synthesis, helping researchers explore a wide range of molecules. The…

  38. LangChain12:29 PMlangchain.com
    Can Jev Be a Better Agent Evaluator?

    We tested using Jev-as-a-Judge against LLM judges on accuracy, repeatability, latency, and cost to see whether System One models could offer a new approach to agent evaluation.

  39. Vercel12:00 PMvercel.com
    Deployments now show billable duration and CPU minutes

    Deployments now show their billable duration and CPU minutes in the dashboard, vc inspect, and the REST API. Use it to understand how each build contributes to your usage. Billable duration is the build and post-build time combined, rounded up to the next whole minute. CPU minute…

  40. GitHub Changelog11:54 AMgithub.blog
    Grok 4.7 is now available in GitHub Copilot

    Grok 4.7, xAI’s latest reasoning model, is now rolling out in GitHub Copilot. Building on Grok 4.6, it is designed for agentic coding and complex, multistep workflows. This model is… The post Grok 4.7 is now available in GitHub Copilot appeared first on The GitHub Blo…

  41. NVIDIA11:51 AMblogs.nvidia.com
    AI Security Is an Engineering Problem — How to Solve It at Every Layer of the Agent Stack

    AI security is an engineering problem. That means defined security requirements, enforceable controls, named owners and evidence that protections work. As AI becomes more capable, the industry must accelerate security engineering, broaden access to defensive tools and share what…

  42. Hugging Face10:44 AMhuggingface.co
    Pruning LLMs Like a Physicist: Block Removal as an Ising Optimization Problem
  43. Cloudflare10:00 AMblog.cloudflare.com
    Python Workers are now generally available

    Python Workers allow developers to run Python web frameworks and AI orchestration libraries natively in the Cloudflare Workers runtime. You can seamlessly integrate with Cloudflare's ecosystem including D1, R2, and Workers AI without writing any JavaScript glue code.

  44. OpenAI9:00 AMopenai.com
    Higgsfield AI ships new video features in a day with GPT-6 Astra

    With GPT-6 Astra, Higgsfield AI makes video ad creation easier for small businesses and brings new creative tools to market faster.

  45. OpenAI9:00 AMopenai.com
    Advisory Group on Mathematics and Artificial Intelligence

    OpenAI is working with an independent Advisory Group on Mathematics and Artificial Intelligence to guide the review and communication of emerging AI results.

  46. Meta Engineering9:00 AMengineering.fb.com
    Inside Petal: Building the World’s First Petabit-Class Transoceanic Subsea Cable

    Petal, the next step in Meta’s subsea innovation, will be the first subsea cable to deliver petabit capacity at transoceanic distances, connecting France and the United States over approximately 7,000 km (4,300 mi). Expected to enter service in 2029, it will be the first subsea c…

  47. NVIDIA7:00 AMblogs.nvidia.com
    5 Companies Using NVIDIA AI for Clean Energy

    Clean energy isn’t hard to come by, but the pace of large-scale adoption has historically been slow due to bottlenecks — including out-of-date infrastructure, elongated research and development timelines, and upfront cost barriers. At New York Climate Week, NVIDIA is highlighting…

  48. OpenAI7:00 AMopenai.com
    Building standards for the next phase of AI

    OpenAI outlines a path to shared global AI standards, calling for coordinated evaluation, reporting, and governance to improve safety.

  49. OpenAI4:00 AMopenai.com
    Expanding OpenAI Academy with new learning paths

    Explore new OpenAI Academy learning paths for employees, developers, leaders, educators, and students to build and demonstrate practical AI skills.

Sunday, September 20

  1. OpenAI9:00 PMopenai.com
    How V7 gives AI agents institutional memory

    Using GPT-5.6, V7 turns scattered company files into context agents can use to complete complex, source-linked work.

  2. Hugging Face9:00 PMhuggingface.co
    tokenizers v1: encode, decode and scaling, measured
  3. xAI9:00 PMx.ai
    Introducing Grok 4.7

    SpaceXAI's most powerful model for coding and knowledge work. Twice as fast, at half the price of comparable models.

  4. Fireworks AI9:00 PMfireworks.ai
    The frontier isn’t a model. It’s a router.

    18 models, 113 coding tasks. The best single model gets 74.1% at $6.52. A perfect router gets 97.6% at $1.88. See how FireRouter closes the gap.

  5. Vercel9:00 PMvercel.com
    MiMo V2.6 models now available on AI Gateway

    MiMo V2.6 Pro, MiMo V2.6 Flash, and MiMo V2.6 Pro UltraSpeed from Xiaomi are now available on AI Gateway. MiMo V2.6 combines coding, reasoning, and tool use with native text, image, audio, and video understanding. Its 1M token context supports long repositories, tool traces, and…

  6. Vercel9:00 PMvercel.com
    AI Gateway now supports TypeSafe clients and an HTTP API for Jev

    You can now call Jev from TypeSafe AI through AI Gateway using an existing TypeSafe client or the HTTP API, in addition to the AI SDK. TypeSafe client: Point an existing TypeSafe client at AI Gateway without changing its evaluation calls. HTTP API: Call Jev directly from any lang…

  7. Vercel9:00 PMvercel.com
    Grok 4.7 now available and 40% off on AI Gateway, fx, and eve

    Grok 4.7 from SpaceXAI is now available on AI Gateway and 40% off through September 27. The discount applies automatically when you call spacexai/grok-4.7. Grok 4.7 has a 500K token context window and supports low, medium, high, and xhigh reasoning levels, giving you control over…

  8. Sourcegraph9:00 PMsourcegraph.com
    The autonomous codebase

    What's left for us to build?

  9. Rust9:00 PMblog.rust-lang.org
    GitHub Actions leaking secrets when Miri output is cached

    The Rust Security Response Team was notified that Miri stores all environment variables to target/, allowing secrets to persist in caches. While not necessary a vulnerability in and of itself, when paired with GitHub Actions caching behavior, it is possible for this to expose sec…

Friday, September 18

  1. Databricks9:00 PMdatabricks.com
    RADAR: Catch gray failures with anomaly detection

    Some of the most damaging outages are the ones your monitoring never flags: a slice...

  2. Vercel5:00 PMvercel.com
    Spend Management expands to Enterprise Flexible Commitment plans

    Enterprise teams on Flexible Commitment plans can now use Spend Management, already available on Pro, at no additional cost. You can set a budget at any time in Spend Management settings. Set a budget per billing cycle, and when your team's metered usage approaches or crosses it,…

  3. Vercel3:00 PMvercel.com
    WebMCP support now available in mcp-handler

    mcp-handler now has experimental support for WebMCP, the proposed web standard for exposing tools to in-browser agents. Add a single script tag to your site, and your existing MCP tools become available there too. Opt tools in by adding them to the experimental_webMcp object: The…

  4. Google Research2:46 PMresearch.google
    MilleMiglia: A realistic instance generator for middle-mile logistics

    Algorithms & Theory

  5. Databricks2:45 PMdatabricks.com
    Enabling secure, productive work on personal devices

    At Databricks IT, our vision is to empower people to work from anywhere without putting...

  6. Cloudflare2:23 PMblog.cloudflare.com
    Saving another 100TB of RAM with math (and Rust)

    Cloudflare's global network is immense but not limitless. As we look for small ways to trim our resource usage, we sometimes get lucky and we can cut significantly more. Here’s how we reduced one of our Pingora-based service's RAM usage with statistics.

  7. Vercel2:00 PMvercel.com
    v0 now reads npm credentials from shared environment variables

    v0 now installs private packages from npm and custom registries using credentials stored as shared environment variables on Vercel. This makes it easier for teams to build with their existing design systems, component libraries, and internal packages directly in v0. To get starte…

  8. Google Cloud1:30 PMcloud.google.com
    Announcing Native BM25 Ranking in AlloyDB and Cloud SQL

    Vector search is a critical component of generative AI, retrieval-augmented generation (RAG), and data agent architectures, but sometimes vector search alone isn't enough. While vector embeddings are incredible at understanding conceptual meaning, they stumble on specific alphanu…

  9. Google Cloud1:00 PMcloud.google.com
    Accelerating the borderless Lakehouse: Announcing preview of cross-cloud caching

    Today, we are excited to announce enhancements to the borderless Lakehouse, our answer to how data engineers, data scientists, and increasingly, AI agents, can query governed data directly where it lives. To reason accurately and automate complex enterprise workflows, agents and…

  10. Google Cloud1:00 PMcloud.google.com
    Reimagining service delivery in the agentic era with Google Public Sector

    State and local governments are driven by a shared mission to provide responsive, equitable, and accessible services. However, achieving this goal is often hindered by legacy technical debt, disconnected data, and heavy administrative burdens that slow down mission delivery. This…

  11. Google Cloud1:00 PMcloud.google.com
    How to upskill enterprise AI builders by using daily micro habits

    As enterprises invest in generative AI, tech leaders keep seeing the same pattern: Developers test AI tools for a week, hit setup problems, and then drift back to the backlog. Nothing ships. The real gap is enablement. In this landmark Harvard Business Review article, Josh Bersin…

  12. Google AI11:00 AMblog.google
    New experts join Google’s AI & Economy team

    We are expanding our AI & Economy team with world-class academic advisors, fellows, and core internal researchers.

  13. JetBrains10:52 AMblog.jetbrains.com
    The AIDEs Framework: How We Built a “Theory of Everything” for AI Development Tools

    We are living in genuinely interesting times. AI is disrupting software development at a pace where new models, tools, and practices appear almost daily. Many teams’ natural first instinct is to spend ever more time chasing updates. After almost two years of AI product and market…

  14. JetBrains10:42 AMblog.jetbrains.com
    Making Local AI Smarter and Faster

    We want local coding agents to be smart and fast, with the ability to understand a codebase, do useful work, and finish tasks without long waits. This Junie Local update makes it practical to use a more capable model on your own machine. In the first release, we had to choose bet…

  15. Google AI10:00 AMblog.google
    Co-creating the future of fashion with Google

    Google worked side-by-side with designers Jane Wade and Sergio Hudson to custom-design Google Flow tools to prep for NYFW.

  16. OpenAI9:00 AMopenai.com
    Introducing the Australian Youth Safety Blueprint

    OpenAI introduces the Australian Youth Safety Blueprint, a six-pillar roadmap for safer AI experiences that protect and empower young people.

  17. Anthropic9:00 AManthropic.com
    Partnering with Accenture on embedded evaluation

    We’re partnering with Accenture on independent evaluation of frontier AI—part of our recent commitment to embed evaluators at Anthropic. Both we and Accenture expect to invest at least $1 billion to build capacity in this area over the next five years.

  18. JetBrains8:51 AMblog.jetbrains.com
    Your diff has a demo now with Junie /demo

    You have the change ready and the tests are green. Now someone has to launch the app, find the right screen, and check the flow. Often, that someone is still you, even when an agent helped write the code. You should be able to delegate that part too. Junie /demo is a new mode in…

  19. JetBrains7:45 AMblog.jetbrains.com
    Ktor 3.6.0 Is Now Available!

    Ktor 3.6.0 is here! This release is full of new experimental features, including typed authentication capabilities with specialized support for OpenID Connect and HTTP/3 support for the Netty engine. There are also a few quality-of-life improvements for routing and request handli…

  20. Databricks6:31 AMdatabricks.com
    Database Branching: A Developer's Guide to Git-Style Workflows

    Git made isolated development a baseline for software teams. Each developer can create...

  21. Vercel4:00 AMvercel.com
    Jev is the fastest-adopted model in AI Gateway history

    Within 24 hours of launching on AI Gateway, Jev from TypeSafe AI reached more than twice as many paid teams as any previous model launch, making it the fastest-adopted model in gateway history. Jev passed every other comparison model in its first twelve hours and continued to wid…

Thursday, September 17

  1. Apple Machine Learning9:00 PMmachinelearning.apple.com
    Dynamically Scaled Activation Steering

    Activation steering has emerged as a powerful method for guiding the behavior of generative models towards desired outcomes such as toxicity mitigation. However, most existing methods apply interventions uniformly across all inputs, degrading model performance when steering is un…

  2. Together AI9:00 PMtogether.ai
    How a global fintech scaled coding agent traffic with Dedicated Model Inference

    Inside a global bank's shift to self-serve dedicated inference: how Together's DMI gave engineering teams direct control over scaling, models, and testing.

  3. xAI9:00 PMx.ai
    Introducing Grok Voice Transcribe 2.0

    Announcing SpaceXAI's newest speech-to-text model, with unparalleled accuracy and cost effectiveness.

  4. Vercel9:00 PMvercel.com
    Reproducing, disclosing, and fixing the libheif vulnerability with Hacktron and the maintainers

    In August 2026, Hacktron reported what looked like a remote code execution (RCE) vulnerability in Next.js image optimization. Their investigation found that the vulnerable code was not in Next.js itself, but upstream in libheif, an AVIF image decoder used by Next.js, ImageMagick,…

  5. Vercel9:00 PMvercel.com
    GLM 5.3 FlashX now available on AI Gateway

    GLM 5.3 FlashX is now available on AI Gateway. GLM 5.3 FlashX is a high-speed serving option for Z.ai's multimodal coding model, delivering inference at ~200 tokens per second for faster streamed responses. The higher serving speed is useful for coding agents, tool loops, and int…

  6. Stripe9:00 PMstripe.com
    Analyzing rising fraud attempts among travel and leisure businesses on Stripe

    Last year, Stripe data shows fraud attempts against travel and leisure businesses hit a four-year high. We analyzed payment activity from more than 200,000 active travel and leisure businesses on Stripe to understand where fraud is rising, how effectively it’s being blocked, and…

  7. GitLab9:00 PMabout.gitlab.com
    Securing the software factory at machine speed

    I joined GitLab at a moment when the way teams build and secure software has been changing rapidly. GitLab CEO Bill Staples recently framed that shift in When Code Is Abundant. When code is no longer the bottleneck, trust becomes scarce, and that constraint shows up first in what…

  8. GitLab9:00 PMdocs.gitlab.com
    GitLab 19.4 released
  9. Vercel7:00 PMvercel.com
    Sub-second artifact deployments are now supported in Vercel CLI

    You and your agents can now deploy static artifacts to Vercel in under one second through Vercel CLI. Run vercel deploy to share a prototype, publish an HTML report, or preview a page created by your coding agent. Vercel automatically detects eligible deployments, and valid artif…

  10. AWS News6:11 PMaws.amazon.com
    New low-cost burstable Amazon EC2 T8i instances are generally available

    AWS introduces new low-cost burstable Amazon EC2 T8i instances powered by custom sixth generation Intel Xeon Scalable Processors (Granite Rapids), available only on AWS. T8i instances are among the lowest-cost EC2 instances and deliver up to 30% better price performance over prev…

  11. Google Research5:45 PMresearch.google
    The future of practice: Enabling teachers to create learning interactives with generative UI

    Education Innovation

  12. Google AI5:00 PMblog.google
    Making global data easier to explore

    Google and the UN system have launched the UN System Data Commons, a new open platform making global statistics accessible and easy to search.

  13. Vercel5:00 PMvercel.com
    Turbo build machines can now be enabled per deployment

    You can now opt into Turbo build machines on any individual deployment. This is useful when you need to increase resources temporarily without changing project settings. You can do this in three ways: Include #VERCEL_BUILD_MACHINE=TURBO in your Git commit message before pushing t…

  14. AWS News4:20 PMaws.amazon.com
    AWS Elastic Beanstalk introduces Cluster Mode

    Run an application on AWS Elastic Beanstalk Cluster Mode without provisioning or operating the compute underneath it. You provide a container image or source code; Elastic Beanstalk with service-operated compute creates and operates the environment that runs it.

  15. Vercel4:00 PMvercel.com
    Run Terminal-Bench and other Harbor evals on Vercel Sandbox

    You can now run Harbor evals on Vercel Sandbox. Harbor is the open-source harness behind Terminal-Bench, whose registry includes many other benchmarks such as SWE-bench, tau3-bench and OSWorld. Pass --env vercel to harbor run and each trial executes in its own isolated Firecracke…

  16. Vercel3:00 PMvercel.com
    The skills CLI now supports Notion hosted skills

    skills@1.7.0 adds Notion skills databases as an install source for agent skills. Notion skills are reusable agent skills written as Notion pages. Teams author, review, and update them in the workspace they already use, then install them into any agent the skills CLI supports. No…

  17. LangChain2:40 PMlangchain.com
    How Included Health Built Federated Healthcare Agents with LangGraph and Deep Agents

    See how Included Health used Deep Agents, LangGraph, and LangSmith to build Dot, a federated healthcare navigation agent with human handoff and clinical oversight.

  18. LangChain2:40 PMlangchain.com
    Building an Agent Harness for Life Sciences: Introducing Deep Life Sci

    Deep Life Sci is LangChain's open source agentic assistant for clinical and lab scientists. It pulls from 600K+ ClinicalTrials.gov studies, 29M PubMed abstracts, and 12M PubMed Central full-text articles, with sandboxed sub-agents for real data analysis.

  19. Databricks2:27 PMdatabricks.com
    Database for AI Agents: 5 Evaluation Criteria

    The five criteria for evaluating a database for AI agents are branch isolation, serverless...

  20. Railway11:59 AMblog.railway.com
    Usage-Based vs. Fixed Pricing: Which Is Cheaper in 2026?

    Provisioned, resource-consumption, and request-based cloud pricing compared, with worked examples showing which model is cheaper for each workload shape.

  21. NVIDIA10:00 AMblogs.nvidia.com
    Cute Critters Come to the Cloud: ‘Aniimo’ Launches on GeForce NOW

    A new creature-catching adventure is ready to stream from the cloud this week. Pawprint Studio’s Aniimo arrives on GeForce NOW at launch, inviting gamers to explore the vibrant continent of Idyll across supported devices. Also this week, 007 First Light receives a path-tracing up…

  22. JetBrains9:39 AMblog.jetbrains.com
    Building a RAG Pipeline for Semantic Code Search: A Developer Diary and Field Notes

    Part 1: Parsing, chunking, and vectorization Some time ago, we set out to build the best semantic code search platform we could: a RAG pipeline that gives LLM agents precise, citable evidence from real repositories instead of whatever grep happens to surface. The eventual solutio…

  23. OpenAI9:00 AMopenai.com
    How Cooley is accelerating IPO work with ChatGPT

    Cooley built GO Public with ChatGPT Work to bring intelligence to the IPO process, helping lawyers surface issues earlier and focus judgment where it matters most.

  24. Anthropic9:00 AManthropic.com
    Introducing the Life Sciences Verification Program

    Anthropic is an AI safety and research company that's working to build reliable, interpretable, and steerable AI systems.

  25. Ai25:00 AMallenai.org
    What a crowdsourced game revealed about steering Olmo 3

    A crowdsourced game built on Olmo 3 showed how people can exploit unexpected model behaviors to stress-test prosocial AI evaluations—and how open access to a model’s internals can help researchers understand why those tests break.

  26. Vercel4:00 AMvercel.com
    Open-weight models take 56% of token volume, Astra doubles Fable 5.1 spend

    AI Gateway Production Index — September 2026 Every month, AI Gateway routes tens of trillions of tokens between production applications and AI labs. That traffic gives us a view of what AI usage actually looks like in today's enterprise, and we publish it here monthly. See the Pr…

Wednesday, September 16

  1. OpenAI9:00 PMopenai.com
    Introducing Astra for Law

    OpenAI for Law brings frontier intelligence for law, custom firm workflows, connected legal data sources, and legal-grade controls for confidential client work.

  2. Apple Machine Learning9:00 PMmachinelearning.apple.com
    REVERSAL-BENCH: A Reversibility Axis and Reset Oracle for Measuring the Reset-Free RL Cliff

    A central goal of autonomous reinforcement learning is continuous policy training without external resets. However, existing paradigms largely depend on underlying environmental reversibility, a property absent in real world manipulation, where events such as pushing objects off…

  3. Weaviate9:00 PMweaviate.io
    4-bit Rotational Quantization

    4-bit Rotational Quantization in Weaviate 1.39: the SIMD performance work, a centered tier, scaling analysis and a TurboQuant comparison.

  4. Fireworks AI9:00 PMfireworks.ai
    Phylo brings frontier AI to more scientists with open models on Fireworks

    Phylo cut inference cost 60% while doubling users month-on-month, running Biomni Lab's long-horizon biology agents on open models on Fireworks.

  5. Vercel9:00 PMvercel.com
    GPT-Live 1 now available on AI Gateway

    GPT-Live 1 from OpenAI is now available on AI Gateway. GPT-Live 1 is a full-duplex voice model and can listen and speak at the same time. Many voice models use turn detection to respond. Full duplex removes that boundary, so a user can pause, interrupt, or add detail while GPT-Li…

  6. Vercel9:00 PMvercel.com
    Native Marketplace integrations now support custom environments

    You can now connect native Marketplace resources to custom environments. Previously, resource connections could only target production, preview, and development environments. Choose custom environments when connecting a resource from the Vercel dashboard, Vercel CLI, or REST API.…

  7. Netlify9:00 PMnetlify.com
    TypeSafe Jev now available in AI Gateway

    TypeSafe’s Jev model is now available through Netlify’s AI Gateway with zero configuration required. Install @typesafe-ai/sdk and use it directly in your Netlify Functions — no API keys to create, no provider config, no base URLs to wire up. AI Gateway handles credentials automat…

  8. PostgreSQL9:00 PMpostgresql.org
    pgAdmin 4 v9.18 Released

    The pgAdmin Development Team is pleased to announce the release of pgAdmin 4 version 9.18. This release of pgAdmin 4 includes 29 bug fixes and new features, including fixes for four security vulnerabilities (CVE-2026-86861 through CVE-2026-86864). For more details, please see the…

  9. Stripe9:00 PMstripe.com
    SaaS platforms are surging despite the SaaSpocalypse

    The SaaSpocalypse was a useful warning for the software industry, but SaaS platforms that help businesses run core operations are more deeply embedded. New platform businesses on Stripe are up 182% year over year.

  10. GitLab9:00 PMabout.gitlab.com
    Rate limits on GitLab.com are changing

    GitLab.com hosts millions of projects for teams of every size that need a platform they can rely on. Demand is climbing quickly, and we expect platform load to grow several times over this year. Predictable limits are what keep GitLab.com fast for everyone on it, including the au…

  11. GitLab9:00 PMabout.gitlab.com
    Optimize your team's price-performance with hosted open weight models

    There’s no single best model for every software development task. Implementing a new feature, diagnosing a failed pipeline, and resolving security vulnerabilities all place different demands on the model handling them. GitLab Duo Agent Platform is expanding GitLab-managed model c…

  12. Rust9:00 PMblog.rust-lang.org
    Be alert: targeted attacks on prominent Rustaceans

    We believe that there is an ongoing campaign targeting rust-lang members and owners of popular crates that is attempting to compromise devices and accounts in order to use them to publish malware. What we've seen A video call is set up for something positive — maybe for a job, ma…

  13. JetBrains5:06 PMblog.jetbrains.com
    New Bug-Fix Releases Are Available for MPS Versions 2026.1.1, 2025.3.2, 2025.2.4, and 2025.1.4

    We’ve released updates for multiple major MPS versions that fix several issues. DOWNLOAD MPS 2026.1.1 Check out all the updates in each particular version below: MPS 2026.1.1 The Projectional Agent Toolkit receives several practical improvements that allow agents to: See an…

  14. Cloudflare5:06 PMblog.cloudflare.com
    When scanners miss the attack: how Cloudflare Client-Side Security protects storefronts

    A modern storefront can look healthy while malicious JavaScript quietly siphons revenue, hijacks clicks, or rewrites analytics. See how Cloudflare's machine learning models surface evasive client-side attacks for analyst investigation.

  15. Kubernetes3:30 PMkubernetes.io
    Kubernetes v1.37: Hardening Container Storage with Bind Mount Options and EmptyDir Permissions

    Kubernetes v1.37 brings important storage security features: emptyDir permission modes and bind mount options. They help application programmers and security professionals implement rigorous security policies, for example, prohibiting deletion of files across containers or execut…

  16. Node.js3:17 PMnodejs.org
    Node.js 26.9.0 (Current)
  17. Vercel3:00 PMvercel.com
    Hobby projects now retain fewer deployments to free up storage

    Hobby projects now retain fewer deployments past the 30-day retention window. Hobby teams get 10GB of Deployment Storage. Every deployment you keep uses some of it, and going over the limit can block you from deploying until you free some up. Deployment Retention for Hobby teams…

  18. AWS News2:50 PMaws.amazon.com
    AWS reimagines the getting started experience

    AWS has reimagined the getting started experience with smart and sensible defaults to help developers get started fast so that they can focus on building. New customers can sign up and get started right away with $100 in Free Tier credits, managed project environments, and simpli…

  19. OpenAI2:00 PMopenai.com
    Our framework for reporting model misalignment

    OpenAI shares a framework for tracking, investigating, and disclosing model misalignment, alongside six reports of unexpected or concerning model behavior.

  20. Vercel2:00 PMvercel.com
    Secure Compute and Static IP builds start 64% faster

    Builds using Secure Compute or Static IPs now start 64% faster, with the average time from deployment creation to build start dropping from 6.7 seconds to 2.4 seconds. Previously, each build waited for a new build container to boot with its network configuration. These builds now…

  21. Vercel2:00 PMvercel.com
    Mem0 joins the Vercel Marketplace

    Mem0 is now available as a native integration on the Vercel Marketplace, giving your AI agents and apps long-term memory. Mem0 remembers user preferences, facts, and context across sessions, so your app stops starting from scratch. Install from the Marketplace with integrated bil…

  22. VS Code2:00 PMcode.visualstudio.com
    Visual Studio Code 1.138

    Learn what is new in Visual Studio Code 1.138 Read the full article

  23. OpenAI1:00 PMopenai.com
    Helping older adults use AI in everyday life

    OpenAI and AARP are bringing free, hands-on ChatGPT workshops to 1,000 older adults across 10 U.S. cities to build practical AI skills safely.

  24. Microsoft Azure1:00 PMazure.microsoft.com
    Microsoft recognized as a Leader in the 2026 Gartner® Magic Quadrant™ for Distributed Hybrid Infrastructure

    Gartner highlighted Microsoft’s unified single-product architecture, flexibility across hyperconverged and disaggregated architectures, and more. The post Microsoft recognized as a Leader in the 2026 Gartner® Magic Quadrant™ for Distributed Hybrid Infrastructure appeared first on…

  25. LangChain12:55 PMlangchain.com
    Scaling Agents in Healthcare & Life Sciences: Lessons from Madrigal Pharmaceuticals, Abridge, and Vizient

    Agent programs in healthcare and life sciences are being built under a different set of constraints than those in most industries. There’s plenty of upside if the constraints can be resolved. Success can mean hours of manual review compressed into minutes, data spread across a do…

  26. NVIDIA12:00 PMblogs.nvidia.com
    NVIDIA Vera Rubin NVL72 Delivers Leading Performance in MLPerf Inference v6.1 Debut

    System performance, efficient infrastructure scaling and continuous software optimization are key levers that determine AI inference economics. Higher system performance means more tokens generated, resulting in higher revenue. Efficient scaling means throughput grows proportiona…

  27. Cohere11:30 AMcohere.com
    Cohere & Aleph Alpha: Transatlantic Sovereign AI

    Cohere and Aleph Alpha partner to launch the first transatlantic sovereign AI solution with dual headquarters in Berlin and Toronto, providing secure and governable AI.

  28. JetBrains11:14 AMblog.jetbrains.com
    Logpoints Walkthrough

    Modern development tools, especially IntelliJ IDEA, have such comprehensive debugging support that for virtually any niche use case, there is a specialized tool for the job. This can make it hard to know where to begin. If you are new to debugging tools and want the biggest retur…

  29. JetBrains11:12 AMblog.jetbrains.com
    Rider and ReSharper 2026.2.2 Are Out!

    We’ve released version 2026.2.2 of ReSharper, Rider, and.NET tools. You can install this update from inside the tools themselves, through the Toolbox App, or on our website. Here’s what’s new in this update. Rider 2026.2.2 AI Agent Setup widget We’ve recently added a wide range o…

  30. NVIDIA10:00 AMblogs.nvidia.com
    Emerald AI, Google and NVIDIA Launch Alliance to Advance Flexible AI Data Centers

    AI factories are the infrastructure of the intelligence era. Scaling them responsibly will depend as much on innovation across the grid as inside the data center. Today, Emerald AI, Google and NVIDIA announced the launch of the AI Energy Management Alliance (AEMA), a first-of-its…

  31. OpenAI10:00 AMopenai.com
    Reimagining advertising with AI

    Explore new AI-powered advertising experiences from OpenAI, including Sponsored Agents, tools for marketers, and integrations with HubSpot and Shopify.

  32. JetBrains9:34 AMblog.jetbrains.com
    Behind the Scenes: How the OpenTelemetry Plugin Maps Your Microservices in Real-Time

    We’ve all been there: you join a new project, and the first thing you ask for is the architecture diagram. You’re handed a diagram that looks great, but after a week of debugging, you realize it’s six months out of date. Service A hasn’t talked to Service B since the…

  33. OpenAI9:00 AMopenai.com
    Hex turns complex analysis into visual reports with GPT‑6 Astra

    GPT-6 Astra helps Hex’s data agents turn answers into interactive visualizations that employees are proud to share.

  34. OpenAI9:00 AMopenai.com
    How to connect AI usage to business value

    Learn how ChatGPT Work and Codex analytics help teams understand AI usage and spend, identify training needs, and connect adoption to business outcomes.

  35. Mistral AI9:00 AMmistral.ai
    Mistral and Mozilla are bringing open, private and multilingual AI to your web browser

    Open, private and multilingual AI is coming to your web browser. Mistral and Mozilla team up to put powerful, trustworthy AI where you already browse.

Columns

The people building AI, writing on their own blogs. Last 30 days.

  1. swyx (Latent Space)Sep 23
    [AINews] Claude Opus 5.5, the new default model for AINews — and everybody cuts prices 40-50%

    overshadowing more efficient GPT6 models from OpenAI

  2. Simon WillisonSep 22
    Claude Opus 5.5, GPT-6 Sol, GPT-6 Luna, and a new price war

    Yesterday was Grok 4.7 ( pelicans ) and MiMo v2.6 Flash/Pro ( more pelicans ). Today Anthropic released Claude Opus 5.5, and around an hour later OpenAI released GPT-6 Sol and GPT-6 Luna. It's going to take a while to get a good read on all of these new models, but here are my im…

  3. swyx (Latent Space)Sep 22
    🔬 An Oscar, Two Asteroids, and the Algorithm in Your sklearn: John Platt on AI for Science

    We talked to Google’s Oscar winning “Giganerd” about automating science, solving climate change, and how future generations can contribute to science in the age of superintelligent AI

  4. Sebastian RaschkaSep 22
    MiMo-V2.6 Pro Architecture and Training Notes

    Notes on MiMo-V2.6 Pro's GQA and sliding-window attention, agent training tasks, reward signals, and large RL batches.

  5. Nathan LambertSep 22
    Debating RSI, the US-China Gap, and Jaggedness with JS Denain of Epoch AI

    Podcast #19

  6. swyx (Latent Space)Sep 22
    [AINews] Xiaomi MiMo-V2.6-Pro 1T-A42B: the new top Open Weights model, trained for $3M

    crowning a new Chinese frontier lab

  7. Simon WillisonSep 21
    Jev introduces a new shape of LLM - System One, aka Decision Models

    Last week TypeSafe AI unveiled Jev, their first example of a new category of model that they are calling "System One models" (I'm with Maggie Appleton, I think "decision models" is a better name for these). Jev is an interesting variant on the usual LLM format: it still accepts t…

  8. Nathan LambertSep 21
    The current balance of power in open models

    The expanded form of a testimony I prepared for Congress.

  9. Sebastian RaschkaSep 20
    It's Easy to Dismiss Jev as Just a Classifier

    A short note on Jev's generalization, possible encoder-style architecture and training, and Choice and Noul API examples.

  10. Thorsten BallSep 20
    Joy & Curiosity #100

    Interesting & joyful things from the previous week

  11. Nathan LambertSep 19
    Why I still haven’t bought into true RSI

    An “AI moderate’s” view on recent events and the trajectory of frontier models.

  12. Ethan MollickSep 18
    The Overhang

    Using your deep knowledge, wide knowledge, taste, and agency

  13. Hamel HusainSep 18
    AI Evals: Everything You Need to Know

    This document curates the most common questions Shreya and I received while teaching 5,000+ engineers and PMs AI Evals. Warning: These are sharp opinions about what works in most cases. They are not universal truths. Use your judgment. How to use this FAQ Browse the questions tha…

  14. Dan AbramovSep 17
    How I Vibed a Proof of Conway’s Conjecture

    You can just prove things, apparently.

  15. Dwarkesh PatelSep 17
    Noam Brown – Agent swarms, alignment, & recursive self-improvement

    “We never want to be in a situation again where we underestimate the AI.”

  16. Martin FowlerSep 17
    I don't like LLMs

    I have a lot of mixed feelings about AI and LLM technology. I’m fascinated by its effect on our profession, excited by the potential gains in productivity - and thus the products we could rapidly build. On the other hand, I’m fearful of the damage AI might cause: agent swarms tak…

  17. Martin FowlerSep 16
    Fragments: September 16

    Reports of agentic hacking continue, in this case it happened back in May and it seems OpenAI did not disclose that they were responsible. Simon Willison sees two options: After the Hugging Face and Wiki attacks OpenAI were still unable to review their previous logs and determine…

  18. Martin FowlerSep 15
    Nail the Narrative

    Sumeet Gayathri Moghe finds many folks building presentations get tangled in building slides without a coherent narrative. He advises distilling the big idea, visualizing the audience, and building a structured storyline. more…

  19. Sebastian RaschkaSep 14
    Pacing != Pacing Development

    My take on AI model pacing as a framework for release checks and the competitive pressure around model releases.

  20. Armin RonacherSep 13
    Interpreting Pangram

    Yesterday David Sacks wrote a tweet and within a few minutes people did, what they usually do, and they asked Pangram if it was AI. And Pangram said it’s entirely AI generated. To which David replied that these AI detectors are bogus. Now Pangram has a pretty low false posi…

  21. Addy OsmaniSep 13
    Brownfield Agentic Engineering

    What it takes to run agents in a codebase older than the team

  22. Simon WillisonSep 12
    Generating running routes with GPT-6 Astra and ChatGPT Work

    Here's a neat thing I had ChatGPT Work with GPT-6 Astra (Max) do this morning: I live at <my address>. Figure out 5K and 10K running routes from me that loop from my house. Use OSM data. It worked for 27 minutes and produced exactly what I'd asked for, as both an embedded v…

  23. Thorsten BallSep 12
    Joy & Curiosity #99

    Interesting & joyful things from the previous week

  24. Simon WillisonSep 11
    OpenAI agents attacked RubyGems back in May

    OpenAI agents carried out an undisclosed attack on RubyGems is a new bombshell report from Spencer Kitts, Thomas Larsen, and Sydney Von Arx - three of the four authors of the report on the agent attack on disused wikis ( previously ) last week. This time they're noting that it lo…

  25. Armin RonacherSep 11
    P(doom)

    This week some flavor of “AI is going to kill us all” went viral. In particular one where an employee put his personal probability of that happening above 10%. Which made me go to the Wikipedia page of P(doom) and I realized that Dario Amodei’s apparent probabil…

  26. Dwarkesh PatelSep 11
    AI researchers debate how close we are to recursive self-improvement

    “We're nowhere near the ceiling.”

  27. Simon WillisonSep 8
    Some thoughts on the Navier–Stokes Millennium Prize Problem

    On the Navier–Stokes Millennium Prize Problem introduces an impressive result from OpenAI, who used an unreleased model to produce a resolution to the Navier–Stokes existence and smoothness problem, one of the seven Millennium Prize Problems that have been subject to a $1,000,000…

  28. Dwarkesh PatelSep 8
    Pretraining progress is mostly coming from data

    Breaking down 6 years of pretraining progress into data vs model improvements

  29. Armin RonacherSep 6
    Astra for Coding: Why Are We Doing This Again?

    I’m more and more convinced that all of AI engineering is Neijuan (内卷, meaning curl inwards). In China it describes a system that demands ever more effort and competition without improving output. The way in which it sometimes shows up in the West is the 996 nonsense. The E…

  30. Thorsten BallSep 6
    Joy & Curiosity #98

    Interesting & joyful things from the previous week

  31. Ethan MollickAug 30
    Agency and Agents

    From the Hugging Face Incident to Twilight Factories

  32. Addy OsmaniAug 30
    Agentic Skill Decay

    Agents can finish the task without teaching you anything. Building expertise now has to be deliberate.

  33. Addy OsmaniAug 26
    Audit your Agent files

    A practical guide to auditing what your coding agent still needs.

Technical Guides

Tutorials, case studies and technical deep dives.

  1. AWS Machine LearningSep 22
    Bring more intelligence to everyday work with GPT-6 Sol and GPT-6 Luna on Amazon Bedrock
  2. AWS Machine LearningSep 22
    Claude Opus 5.5 is now available on AWS
  3. NVIDIA DeveloperSep 22
    Enabling Private High-Performance Production AI Inference with NVIDIA Confidential Computing
  4. AWS Machine LearningSep 22
    Evaluate skill-equipped agents with Strands Evals and Amazon Bedrock AgentCore
  5. NVIDIA DeveloperSep 22
    Topology-Aware Workload Scheduling with NVIDIA Topograph
  6. AWS Machine LearningSep 22
    How Reactiv automates mobile commerce 80% faster with Amazon Bedrock AgentCore
  7. AWS Machine LearningSep 22
    Right-size generative AI endpoints with concurrency sweeps on Amazon SageMaker AI
  8. AWS Machine LearningSep 22
    How Trane gets building insights 60x faster with Amazon Bedrock AgentCore
  9. AWS Machine LearningSep 22
    How Tata Elxsi detects industrial safety risks in seconds on AWS
  10. AWS Machine LearningSep 22
    Extending public sector intelligence with Agentforce and AWS
  11. NVIDIA DeveloperSep 22
    What’s New for Game Developers: DLSS 5 with 3D-Guided Neural Rendering, NVIDIA ACE Updates, and New RTX Kit Capabilities
  12. NVIDIA DeveloperSep 22
    Accelerating a ROS 2 Node with an AI Agent and NVIDIA Isaac ROS
  13. NVIDIA DeveloperSep 21
    Simplifying Model Serving Across Multiple GPUs with NVIDIA TensorRT Multi-Device Integration in NVIDIA Dynamo-Triton
  14. NVIDIA DeveloperSep 21
    How to Evaluate AI Agents From Tool Calls to Task Completion
  15. AWS Machine LearningSep 21
    xAI’s Grok 4.6 is now available in Amazon Bedrock
  16. AWS Machine LearningSep 21
    How BMW Group detects cost anomalies across 14,000 cloud accounts
  17. AWS Machine LearningSep 21
    Run Positron on Amazon SageMaker AI for data science workflows
  18. AWS Machine LearningSep 21
    How Benchling secured multi-tenant AI agents with Amazon Bedrock AgentCore
  19. AWS Machine LearningSep 21
    Reducing medical claims review time with AI on AWS: The EXL Medical IDP solution
  20. NVIDIA DeveloperSep 21
    Turn Your Latest Observations Into Timely Weather Decisions With NVIDIA Earth-2
  21. AWS Machine LearningSep 18
    Amazon SageMaker Inference: 2026 year-to-date launches in review
  22. NVIDIA DeveloperSep 18
    Benchmarking LLM Inference at Scale with AIPerf
  23. AWS Machine LearningSep 18
    Introducing Kimi K3 on Amazon Bedrock
  24. AWS Machine LearningSep 18
    Migrating multi-model AI agents to Amazon Bedrock AgentCore runtime
  25. AWS Machine LearningSep 18
    The new AgentCore runtime: Elastic, optimized, and consistently fast starts
  26. AWS Machine LearningSep 18
    Deploy Hugging Face models on Amazon SageMaker AI with coding agents
  27. AWS Machine LearningSep 18
    Introducing Amazon SageMaker HyperPod Inference Gateway
  28. AWS Machine LearningSep 17
    Reduce time-to-hire for quality candidates with AI-powered Amazon Connect Talent
  29. NVIDIA DeveloperSep 16
    How to Use AI Agents to Prepare 3D Scenes for Simulation
  30. NVIDIA DeveloperSep 16
    TensorRT Edge-LLM Completes the MLPerf Edge Agentic Benchmark 6.4x Faster on Jetson AGX Thor
  31. NVIDIA DeveloperSep 16
    Translating CUDA Tile Operations from Python to Rust Using Agentic AI

New Models

Everything that joined the OpenRouter catalog.

  1. OpenRouterSep 23
    New model: Upstage: Solar Mini 4
  2. OpenRouterSep 22
    New model: Cohere: Command A+
  3. OpenRouterSep 22
    New model: OpenAI: GPT-6 Luna Pro
  4. OpenRouterSep 22
    New model: OpenAI: GPT-6 Luna
  5. OpenRouterSep 22
    New model: OpenAI: GPT-6 Sol Pro
  6. OpenRouterSep 22
    New model: OpenAI: GPT-6 Sol
  7. OpenRouterSep 22
    New model: inclusionAI: Ming Image 0.1 Design
  8. OpenRouterSep 22
    New model: Anthropic: Claude Opus 5.5
  9. OpenRouterSep 22
    New model: AssemblyAI: Universal-3.5 Pro
  10. OpenRouterSep 21
    New model: Xiaomi: MiMo-V2.6-Pro-UltraSpeed
  11. OpenRouterSep 21
    New model: Xiaomi: MiMo-V2.6-Flash
  12. OpenRouterSep 21
    New model: Xiaomi: MiMo-V2.6-Pro
  13. OpenRouterSep 21
    New model: SpaceXAI: Grok 4.7
  14. OpenRouterSep 21
    New model: Qwen: Qwen3.8 Omni Flash
  15. OpenRouterSep 18
    New model: PrismML: Ternary Bonsai 2 27B
  16. OpenRouterSep 18
    New model: Z.ai: GLM 5.3 FlashX
  17. OpenRouterSep 17
    New model: TypeSafe: Jev Latest
  18. OpenRouterSep 17
    New model: TypeSafe: Jev 1.13
  19. OpenRouterSep 17
    New model: Pareto

Research

New arXiv papers in AI, language, machine learning and software engineering. The 30 newest; with Jev on, the 30 most relevant.

  1. arXiv cs.CLSep 22
    Flash-dLLM: IO-Aware KV Caching and Parallel Decoding for Fast, Memory-Efficient Diffusion LLMs

    Diffusion Large Language Models (dLLMs) have recently emerged as a promising alternative to autoregressive LLMs by enabling non-autoregressive text generation. However, their practical deployment remains limited by inefficient inference, largely due to the absence of effective Ke…

  2. arXiv math.OCSep 22
    A Decentralized Partially Observable Team Decision Methodology with Delayed Information Sharing

    We study decentralized partially observable team decision problems with low-rank latent dynamics and unknown system models. The proposed framework combines team-theoretic equivalence with low-rank model representations to address cooperative decision-making in partially observabl…

  3. arXiv cs.CLSep 22
    Agensh: Scaling Organizational Intelligence to 1,024 Agents

    A multi-agent system can reduce latency on complex tasks by executing work concurrently. Several pioneering harness frameworks support multi-agent systems. However, the scalability of current multi-agent harnesses is often constrained by a central orchestrator's capacity to alloc…

  4. arXiv cs.CLSep 22
    SpeakerMem-R1: Speaker-Centered Dual-Track Memory for Multi-Party Dialogue

    Long-term conversational memory in multi-party settings requires more than retrieving relevant content from long-term conversations: it must distinguish who said what, whom each statement concerns, how individuals perceive one another, what information is shared by the group, and…

  5. arXiv cs.AISep 22
    CliffCompaction: Cost-Efficient Compaction for Long-Horizon Coding Agents

    Agents often work on complex problems that require millions of tokens of context, which necessitates compacting across sessions due to limited context windows. We develop CliffCompaction, an autocompaction technique that reduces cost by up to 50% under a bounded context while mai…

  6. arXiv cs.AISep 22
    SWE-Serve: Benchmarking Agentic Engineering For Production Inference Serving

    We introduce SWE-Serve, a benchmark for evaluating agents on production inference engineering tasks. Implementing an inference feature can require coordinating multiple changes across the serving stack, including model support, runtime execution, and public APIs. Existing benchma…

  7. arXiv cs.CRSep 22
    A2M: Trace-Optimized Agent Hijacking in the MCP Ecosystem

    Agents using the Model Context Protocol (MCP) rely on semantic matching to select tools from third-party servers, exposing a semantic supply-chain risk through attacker-controlled metadata and outputs. We introduce A2M (Attraction-to-Manipulation), a two-stage black-box framework…

  8. arXiv cs.AISep 22
    Grow the Harness, Not the Context: From Strategy-Free Scaffolds to Reusable Specialist Agents

    Large language model (LLM) agents often handle streams of related tasks, yet standard harnesses repeatedly ask the model to reconstruct the same control decisions inside each task's context. We study whether task feedback can instead turn recurring control into reusable executabl…

  9. arXiv cs.AISep 22
    Type-Safe Is Not Error-Free: A Constrained Decision Head Follows the Option Name, Not the Rubric Bound to It

    Typed decision models are built for settings where model outputs are consumed directly by software. Instead of generating free-form text, they return a decision over a predefined set of options. By construction, every output conforms to the required schema. Yet this guarantee doe…

  10. arXiv cs.CVSep 22
    FleXray: Universal Clinical X-ray Segmentation

    X-ray is medicine's most widely used imaging modality, yet remains among its least quantitative. Unlike volumetric modalities like CT or MRI, X-ray collapses 3D anatomy into a 2D projection, causing structures to overlap and anatomical boundaries to be ambiguous, even to experts.…

  11. arXiv cs.LGSep 22
    EquivSVA: A Formally Verified Dataset of Behavioral Assertions Across Equivalent RTL Implementations

    Large language models are increasingly used to generate SystemVerilog Assertions from natural-language specifica- tions and register-transfer-level designs. Existing datasets and benchmarks support important goals such as large- scale training, formal evaluation, specification-to…

  12. arXiv cs.SESep 22
    Metrics Failure in LLM-Based Code Vulnerability Repair: An Empirical Study and a Change-Aware Screen

    Large language models (LLMs) are increasingly applied to the automated repair of C/C++ security vulnerabilities, and compile rate is a commonly reported proxy for progress: whether the generated patch compiles. We argue that compile rate is a scientifically unreliable metric for…

  13. arXiv stat.MESep 22
    Automatic depth-based local center clustering via $β$-integrated local depth and adaptive grouping

    Clustering is an unsupervised learning technique that partitions unlabeled data into groups. Most existing methods require user-specified parameters, such as the number of clusters or neighborhood size. Conversely, we propose automatic depth-based local center clustering (A-DLCC)…

  14. arXiv cs.SISep 22
    Diffusion-Induced Spatial Attention Overlapping Community Detection

    Detection of overlapping communities is essential for modelling networks in which nodes participate simultaneously in multiple structural or functional groups. Existing graph neural network approaches commonly rely on local message passing, which can obscure community boundaries…

  15. arXiv cs.HCSep 22
    Does AI Save Time on Product Design? A Randomized Controlled Experiment of AI Prompt-to-Design Workflows

    AI tools for digital product design now offer prompt-to-design capabilities, allowing designers and their non-designer colleagues to create prototypes through conversational workflows with large language models (LLMs). While these tools promise time savings, experimental evidence…

  16. arXiv cs.LGSep 22
    The Sirens' Song: When Proximal Background Context Overshadows Distant Evidence

    Long-context LLMs focus on retrieving distant evidence from extensive context, yet existing work has largely focused on overcoming distance alone. In this work, we identify the Proximity Trap, insufficient attention to distant evidence often arises less from distance itself than…

  17. arXiv cs.SESep 22
    TraceVIC: Causal Reasoning over Code Evolution for Identifying Vulnerability-Inducing Commits

    Software vulnerabilities are often discovered long after they are introduced, making it difficult to identify the vulnerability-inducing commit (VIC) responsible for introducing the underlying vulnerable condition. Existing VIC identification techniques largely rely on git blame…

  18. arXiv cs.LGSep 22
    Train Where the Quantized Model Goes: On-Policy Distillation for Low-Bit Reasoning

    Quantization-aware distillation (QAD) restores much of the short-form question-answering performance lost to sub-3-bit quantization, yet leaves mathematical and code reasoning substantially impaired. Long generations often degenerate into repetitive loops, exhausting the decoding…

  19. arXiv stat.MESep 22
    Optimal Sequential Annotations for Off-Policy Evaluation

    Offline reinforcement learning and off-policy evaluation evaluates dynamic treatment rules based on retrospectively collected data prior to deployment. In recent AI applications, state and reward information is recorded as complex text or image, which recent AI advancements such…

  20. arXiv quant-phSep 22
    When are bosonic Gaussian states classical to learn?

    A fundamental question in physics is: When does classical behavior emerge from quantum systems? Bosonic Gaussian states provide a natural setting to explore this quantum-classical boundary, as they capture both the classical field behavior and the intrinsic quantum nature of ligh…

  21. arXiv cs.CLSep 22
    Beyond Repeated Sampling: Learning Search Policies for LLM Reasoning

    Large language models increasingly tackle hard reasoning problems by spending more test-time compute, yet the dominant strategy remains naive repeated sampling: draw many independent solutions and hope one is correct. Because such sampling explores only through local decoding noi…

  22. arXiv cs.CLSep 22
    Measuring the Serving Stack Instead of the Model: Hidden Confounds in Local Tool-Use Evaluation

    A coding agent must emit a valid tool call--a parseable invocation of a tool in the provided schema--before the harness can execute its chosen action. We study how local serving stacks affect this protocol step and show that measured outcomes can depend on the serving layer rathe…

  23. arXiv cs.CLSep 22
    Detecting GPT-Assisted Writing Using Interpretable Stylometric Features

    Distinguishing GPT-assisted from independently authored student writing has become a critical challenge in academia. This paper evaluates the discriminative capability of interpretable stylometric features extracted solely from submitted text. Using data from 90 participants who…

  24. arXiv astro-ph.SRSep 22
    PROSWIN: Probabilistic Solar Wind Speed Forecasting Using Deep Distributional Regression From Solar Images

    Accurately predicting fast solar wind conditions is challenging, as uncertainties are large and unquantified by traditional single-value prediction models. In particular, the risks of high-speed solar wind streams (HSSs), which can cause damage to technological infrastructure, ca…

  25. arXiv cs.CRSep 22
    From Alignment to Access Control: A Framework for GenAI Policy Enforcement

    Generative AI (GenAI) applications have flourished enabling users to chat with large language models, and to create agents to act on their behalf for a variety of tasks. The pace of development of capabilities in this field is incredibly fast with security and safety taking a bac…

  26. arXiv cs.LGSep 22
    A Spectral Theory of Grokking: Weight Decay induces Feature Learning

    In grokking an early fit to the training data separates from a much later improvement in generalization. During this delay, training can move from a fixed neural tangent kernel (NTK) regime to one in which task-relevant kernel eigendirections continue to evolve. We provide a quan…

  27. arXiv cs.LGSep 22
    MAGIC: Mixed-Granularity Agent Graphs via Incremental Construction with Dense-Reward Reinforcement Learning

    Collaboration topology shapes both the performance and execution cost of LLM-based multi-agent systems. Because tasks differ in complexity and required capabilities, recent approaches generate task-specific collaboration graphs that specify agent participation and information flo…

  28. arXiv cs.IRSep 22
    Discovery-Driven Integration of Disjoint Tables via Text

    Integrating heterogeneous datasets within data lakes is a critical challenge, particularly for semantically related tables that lack the explicit attributes needed to be joined. We study Discovery-Driven Integration, where the relevant sources and their missing relational structu…

  29. arXiv math.STSep 22
    Statistical Rates for Entropic Optimal Transport in the Discrete to SubGaussian Regime

    We study statistical rates in entropic optimal transport in the semi-discrete regime where one measure has finite support and the other is subGaussian. Our main result establishes parametric convergence rates for the empirical dual potentials to their population counterparts, wit…

  30. arXiv cs.AISep 22
    The Delegation Blind Spot: Auditing Product Decisions from Agent Choices

    Successful agent execution need not identify which future product improvement its user would value. We present a decision-specific audit that maps a declared observation channel and product-value contrast to compatible intervals and witness populations. Its foundations are establ…

Releases

Stable versions of the tools and libraries we track. Prereleases are left out.

ProjectVersionDate
llama.cpp b11125 Sep 23
llama.cpp b11124 Sep 23
llama.cpp b11123 Sep 23
ai 7.0.112 Sep 23
vLLM v0.30.1rc0: [ROCm][CI] Add MI355 dense NVFP4 and MoRI kernel mirrors (#58281) Sep 23
openai 7.23.0 Sep 23
Zed nightly Sep 23
OpenAI Codex 0.156.1 Sep 22
Ollama v0.34.4 Sep 22
LangChain langchain-openai==1.6.4 Sep 22
LangChain langchain-anthropic==1.7.3 Sep 22
@openrouter/sdk 1.3.19 Sep 22
Ollama v0.34.3 Sep 22
Next.js v15.5.26 Sep 22
Next.js v16.3.6 Sep 22
@anthropic-ai/sdk 0.128.0 Sep 22
Claude Code v2.1.280 Sep 22
next 16.3.6 Sep 22
Zed collab-staging: Add Grok 4.7 to Supergrok and XAi (#64567) Sep 22
vLLM v0.30.0 Sep 22
@google/genai 2.24.0 Sep 22
LangChain langchain-fireworks==1.6.2 Sep 22
LangChain langchain-deepseek==1.1.1 Sep 22
LangChain langchain-openrouter==0.2.9 Sep 22
langchain 1.5.12 Sep 21
LangChain langchain-openai==1.6.3 Sep 21
LangChain langchain-core==1.6.4 Sep 21
LangChain langchain-typesafe==0.0.1a3 Sep 20
Claude Code v2.1.278 Sep 19
vLLM v0.30.0rc2 Sep 18
Claude Code v2.1.277 Sep 18
LangChain langchain==1.4.2 Sep 18
Claude Code v2.1.276 Sep 17
Ollama v0.34.2 Sep 17
Claude Code v2.1.275 Sep 17
LangChain langchain-typesafe==0.0.1a2 Sep 17
vLLM v0.30.0rc1: [Bugfix] Isolate supplemental FlashInfer BF16 autotuning (#57285) Sep 17
Zed html-v0.3.2: html: Bump to v0.3.2 (#64371) Sep 17
Zed glsl-v0.2.5: glsl: Bump to v0.2.5 (#64370) Sep 17
Zed proto-v0.3.4: proto: Bump to v0.3.4 (#64372) Sep 17
Deno v2.9.7 Sep 17
Zed v1.20.2 Sep 17
vLLM proto-v0.3.0 Sep 17
Claude Code v2.1.274 Sep 16
Zed v1.20.1 Sep 16

New Repos & Packages

New repositories under the GitHub topics we track, and packages published to npm.

  1. npm #llmSep 23
    gitlab-ai-provider 6.18.0

    GitLab Duo provider for Vercel AI SDK

  2. npm #mcpSep 23
    n8n-mcp 2.88.0

    Integration between n8n workflow automation and Model Context Protocol (MCP)

  3. npm #ai-agentSep 23
    instar 1.3.1259

    Coherence infrastructure for self-evolving AI agents — on the Claude Code or Codex subscription you already have.

  4. npm #ai-agentSep 23
    @xmanrui/dsh-im 4.27.0

    把十二种 IM 渠道和公网 AI Office 接入本机 DeepSeek Harness。 Connect twelve IM channels and a public AI Office to a local DeepSeek Harness.

  5. npm #mcpSep 23
    @lovable.dev/mcp-js 3.0.2

    Author MCP servers for Lovable apps. Declare tools with defineTool, register them in defineMcp, and a framework adapter (TanStack or Supabase Edge Functions) emits the route(s) at build time.

  6. npm #llmSep 23
    @mastra/core 1.69.0

    Mastra is a framework for building AI-powered applications and agents with a modern TypeScript stack.

  7. npm #llmSep 23
    mastra 1.31.1

    cli for mastra

  8. npm #mcpSep 23
    hostinger-api-mcp 1.63.3

    MCP server for Hostinger API

  9. npm #mcpSep 23
    @ai-sdk/mcp 2.0.56

    The **Model Context Protocol (MCP) client** for the [AI SDK](https://ai-sdk.dev/docs) lets you connect to MCP servers and use their tools with AI SDK functions like `generateText` and `streamText`.

  10. npm #mcpSep 23
    pi-mcp-adapter 2.37.0

    MCP (Model Context Protocol) adapter extension for Pi coding agent

  11. npm #ai-agentSep 22
    agent-afk 5.230.9

    Open-source coding-agent harness you can actually change — own the loop (prompts, gates, routing, skills, terminal states), use any model, run long tasks while you're away.

  12. npm #ai-agentSep 22
    @askjo/camofox-browser 1.17.0

    Headless browser automation server and OpenClaw plugin for AI agents - anti-detection, element refs, and session isolation

  13. npm #mcpSep 22
    eve 0.64.1

    Filesystem-first framework for durable backend AI agents that run anywhere.

  14. npm #ai-agentSep 22
    @plannotator/pi-extension 0.27.18

    Plannotator Pi extension - interactive plan review with annotations, annotate agent messages, and review code/PRs

  15. npm #ai-agentSep 22
    @zereight/mcp-gitlab 2.1.65

    GitLab MCP server for projects, merge requests, issues, pipelines, wiki, releases, and more

  16. npm #mcpSep 22
    @copilotkit/aimock 1.43.0

    Mock infrastructure for AI application testing — point your SDK at one local port and every provider, protocol, and service answers deterministically.

  17. npm #llmSep 22
    @earendil-works/pi-coding-agent 0.87.1

    Coding agent CLI with read, bash, edit, write tools and session management

  18. npm #llmSep 22
    @earendil-works/pi-ai 0.87.1

    Unified LLM API with automatic model discovery and provider configuration

  19. npm #llmSep 22
    @earendil-works/pi-agent-core 0.87.1

    General-purpose agent with transport abstraction, state management, and attachment support

  20. npm #mcpSep 22
    agent-device 0.21.12

    Mobile app automation and verification for AI coding agents. CLI, MCP server, and typed Node.js API for iOS, Android, HarmonyOS, TV, web, macOS, and Linux.

  21. npm #ai-agentSep 22
    @humanspeak/svelte-markdown 1.9.2

    Markdown and HTML renderer for Svelte 5 — built for rendering streaming AI agent output from Claude Code, ChatGPT, and agentic workflows. XSS-safe defaults, streaming-aware sanitization, token caching, TypeScript types, and Svelte 5 runes.

  22. npm #ai-agentSep 22
    @mastra/github-signals 0.5.0

    GitHub PR notification signal provider for Mastra agents

  23. npm #llmSep 22
    typia 15.0.0

    Superfast runtime validators with only one line

  24. npm #llmSep 22
    thinking-orbs 0.3.2

    Dotted thought-orb loading indicators for AI & agent UIs — nine tuned states, two sizes, auto dark/light

  25. GitHub #ai-agentsSep 21
    kydlikebtc/awesome-jev (★ 169)

    805 verified examples of Jev — TypeSafe AI's System One decision model — indexed by the decision each one makes, not the blog that mentioned it. Every cited call site is re-read by CI each week. Bilingual EN/中文, JSON schema, and a cross-platform compatibility table.

  26. GitHub #generative-aiSep 21
    scarletkc/Variora (★ 11)

    Different models, the same brief - a collection of demos built from shared prompts, with source, screenshots, and notes.

  27. GitHub #llmSep 21
    jev-chat/jev-chat-windows (★ 398)

    微信(Windows 4.x)旁挂的回复辅助:窗口截图 + 本地离线 OCR 读对方消息 → Jev 判断意图 → 3 条候选一键填入,发送永远手动

  28. GitHub #llmSep 21
    jev-chat/jev-chat-jarvis-mac (★ 290)

    微信消息意图识别悬浮窗(macOS):看屏 + 本地小模型判断意图和风险,再按话术生成回复候选。纯只读、不注入微信。

  29. GitHub #llmSep 21
    jev-chat/jev-chat-jarvis (★ 4,967)

    装在手机上的对话副驾:在微信 / QQ / X / 飞书里读懂对方、给出候选回复、一键填入输入框,发不发由你。非侵入,只读屏幕,不 hook 不改包。

  30. GitHub #llmSep 21
    nokia-applied-research/AnyJev (★ 282)

    Turn any LLM into a Jev-style decision model: typed decisions, real probabilities, no training. (continue updating)

  31. GitHub #generative-aiSep 20
    Rylaispirit/cinematic-video-prompt-skill (★ 71)

    AI video prompt cheat sheet & Claude Skill: cinematic camera angles, camera movement, lighting, composition, color grading for Veo 3, Kling, Sora, Runway, Midjourney. 700+ terms with Vietnamese explanations.

  32. GitHub #generative-aiSep 20
    Inkloom-art/inkloom (★ 62)

    Specialised AI models for logo design — a brand-analysis model turns a business into constraints, typography and symbol models construct the mark, and a composition engine produces real lockups and clear-space rules. Early access open.

  33. GitHub #generative-aiSep 20
    Anil-matcha/awesome-agent-apis (★ 6)

    660+ muapi-hosted generative-media models plus community-submitted third-party API tools (SEO, enrichment, social, scraping) — one YAML file per entry, browsable by capability.

  34. GitHub #mcpSep 19
    nassim-arifette/jevgrep (★ 57)

    Jev-powered semantic code search for coding agents — find behavior across repositories via CLI or MCP, with exact source excerpts and line numbers.

  35. GitHub #mcpSep 19
    walidboulanouar/awesome-jev-use-cases (★ 121)

    Awesome list of TypeSafe AI Jev use cases: 74 demos ranked by likes, 150+ GitHub repos, limits, cost and API examples. CC0, sponsored by AY Automate.

  36. GitHub #llmSep 19
    v-modal/awesome-jev-tools (★ 676)

    A curated list of tools built for Jev — TypeSafe AI's System One model for typed decisions.

  37. GitHub #mcpSep 19
    hirotomasato/yowes (★ 164)

    Generate realistic teacher documents (ID cards, licenses, letters) for 13 countries via MCP.

  38. GitHub #mcpSep 18
    kraayenjon/awesome-jev (★ 116)

    A curated list of Jev use cases, projects, SDKs, and resources. Jev is TypeSafe AI's System One model for fast, typed decisions in software — Choice, Score, and Noul with calibrated probabilities.

  39. GitHub #ai-agentsSep 18
    cobanov/awesome-jev (★ 362)

    A curated, source-backed list of projects built with Jev, TypeSafe AI's System One model for typed decisions.

  40. GitHub #ai-agentsSep 18
    Oldcircle/geo-sleuth (★ 285)

    An agent skill that finds where a photo was taken — OpenStreetMap geometry, elevation skylines, satellite imagery and street view — and shows its work. Works with Claude Code, Codex, Cursor, Gemini CLI, OpenCode and GitHub Copilot.

  41. GitHub #generative-aiSep 18
    joeseesun/bilibili-ai-arena (★ 12)

    中国创作者给全球 AI 出的一场真实大考 | A searchable index of Bilibili AI Infinite Arena videos.

  42. GitHub #ai-agentsSep 18
    logicrw/awesome-jev-projects (★ 416)

    Awesome Jev: source-backed open-source ecosystem radar, plain-language project discovery, and automatic GitHub sync

  43. GitHub #llmSep 17
    ekzhang/openjev-sglang (★ 288)

    Jev-compatible API endpoint based on open models (prefill-only)

  44. GitHub #developer-toolsSep 17
    repoboost-hq/github-launch-checklist (★ 194)

    Audit a GitHub repository before launch - ten readiness checks scored 0-10: name, description, topics, README, license, demo, installation, contributing guide and issue templates. Free and open source.

  45. GitHub #llmSep 17
    yibie/awesome-jev (★ 1,408)

    A curated list of public projects, integrations, and discussions built on Jev — TypeSafe AI's System One model for typed decisions.

  46. GitHub #generative-aiSep 17
    inikolax/remiqora (★ 119)

    Local AI music studio unifying ACE-Step 1.5 and YuE2-3B in one Vue interface — text-to-music generation, stem separation, MIDI transcription, and LoRA fine-tuning, with a built-in multitrack DAW.

  47. GitHub #llmSep 17
    AbdelStark/awesome-typesafe-jev (★ 487)

    Awesome Jev: a source-backed field guide to TypeSafe's System One model, with SDKs, live demos, agent tools, and independent evaluations.

  48. GitHub #developer-toolsSep 17
    nMaas8388/github-ranking-audit (★ 212)

    Audit your GitHub repository search ranking signals. Checks name, description, topics, README, stars, forks, and activity.

  49. GitHub #ai-agentsSep 17
    ethanplusai/astra-flash-orchestrator (★ 612)

    Astra plans and reviews; DeepSeek Flash builds. A native Codex workflow with phased tasks, verification, safe installation and reversible setup.

  50. GitHub #ai-agentsSep 17
    NiazMorshed2007/jev-review (★ 211)

    Local-first MCP plugin for continuous software-quality review by AI coding agents, powered by Jev.

  51. GitHub #mcpSep 17
    itsmostafa/typesafe-mcp (★ 270)

    An mcp connector to evaluate anything fast and cheap. Give your AI agent direct access to typesafe ai's jev model.

  52. GitHub #mcpSep 16
    kajeesan/Open-Health-Atlas (★ 60)

    Explore sleep, training, mood and nutrition together using local records, traceable calculations and optional MCP tools for your preferred AI client.

  53. GitHub #ai-agentsSep 16
    awlevin/typesafe-computer-use (★ 863)

    Computer use for about $0.0002 a step: OCR the screen, classify the next action with TypeSafe, click. macOS.

  54. GitHub #developer-toolsSep 16
    Marcos66236/github-stars-history (★ 277)

    Track and visualize the stars history of any GitHub repository. Open-source growth analytics and velocity tracking.

  55. GitHub #mcpSep 16
    Ying-Kai-Liao/jev-browser (★ 76)

    Browser automation where an LLM plans and Jev (Typesafe System One) decides. Library, CLI and MCP server.

  56. GitHub #generative-aiSep 16
    cporter202/ai-apis-you-can-ship-today (★ 14)

    This GitHub repo is a powerhouse collection of AI APIs you can start using immediately to ship real products — from LLM tools to computer vision and agents. One of the most valuable AI API lists on GitHub.

Open at the source ↗