Press "Enter" to skip to content

the boring bot

Anthropic announces Claude Fable 5.1 and Claude Mythos 5.1

Anthropic said in their announcement, “Claude Fable 5.1 and Claude Mythos 5.1 are the same model, but with different levels of safeguards. Fable 5.1 is generally available, while Mythos 5.1 is available only through our trusted access programs; its safeguards are specifically designed to support work in cybersecurity and the life sciences. Alongside its increased capabilities, Fable 5.1 takes important steps towards addressing the feedback we’ve received from customers on price, data retention, and safeguards.”

May be related to Anthropic’s previous announcement, “Fable 5.1 will cost an estimated 25% less than Fable 5 for typical workloads, wherever usage is billed by token. This is because we’re reducing our pricing on cache reads (where the model reads inputs that have already been processed and stored). For highly agentic work, the savings will often be much larger—up to approximately 45%.”

Key Upgrades & Technical Performance

  • Next-Level Coding & Benchmarks: Fable 5.1 sets new performance records, scoring 52.6% on the Terminal-Bench-Science 0.1 benchmark (compared to 24.7% for Fable 5) and reaching 73.4% on CursorBench 3.2.0. The model excels at finding root causes in complex codebases—in one instance identifying a rare bug in an external vendor library that had eluded human engineers for years.
  • Mythos 5.1 for Specialized Research: Available strictly through trusted access programs, Mythos 5.1 utilizes specialized, high-tier safeguards designed for advanced cybersecurity vulnerability discovery and scientific research in the life sciences.
  • Significant Cost Reduction: Typical token-based workloads will see an estimated 25% cost reduction compared to Fable 5 due to lower pricing on cache reads. Highly agentic workflows can yield total savings of up to 45%.
  • Enterprise Frontier Safeguards (EFS): Anthropic introduced EFS to deliver zero data retention capabilities by allowing enterprise users to store data directly within their own cloud infrastructure rather than on Anthropic servers.

Benchmark Comparison

Metric / Benchmark Claude Fable 5.1 Claude Fable 5 GPT-5.6 Sol
Agentic Coding (Terminal-Bench 4.0) 55.8% (60.9% for Mythos) 42.0% 37.3%
Scientific Research (Terminal-Bench-Science 0.1) 52.6% 24.7% 22.4%
Multidisciplinary Reasoning (Humanity’s Last Exam – with tools) 65.0% 63.8%

ICYMI: Google Launches AI-Powered Shopping Features in Gemini for the Philippines

Google has expanded its e-commerce capabilities by rolling out AI-driven shopping features within the Gemini app in the Philippines. The update allows local users to research, compare products, check prices, and find where to purchase items—all inside a single conversational chat.

Google announced the new shopping feature in Gemini late last June.

Manus AI Returns to Independent Control After Meta Deal Unwinds

Manus, the AI agent startup with Chinese roots, is reverting to independent operation after its acquisition by Meta Platforms fell apart. Meta had agreed last December to buy Manus for more than $2 billion, planning to fold its autonomous “general-purpose” agent technology into products like Meta AI and WhatsApp. But in April, Chinese regulators ordered the deal reversed as Beijing tightened scrutiny of U.S. investment in Chinese firms developing frontier technology.

Announcing the transition, Manus said its founding team — Xiao Hong, Zhang Tao, and Ji Yichao — will continue leading the company, “with a relentless commitment to product innovation and advancing general AI agents for our users worldwide.” As an independent lab, Manus said users can expect it to become more embedded in daily workflows, interact more directly with the world, and act more proactively on users’ behalf. The company also thanked those who backed up their data during the transition.

Tencent has reportedly been in talks to become Manus’ largest shareholder.

Anthropic to Adjust Claude Code Limits: 25% Permanent Increase Replaces 50% Promo on Sept. 14

On the official Claude Dev X/Twitter account, the company announced: “Starting September 14, we’re permanently raising standard weekly limits in Claude Code by 25% for Pro, Max, Team, and seat-based Enterprise plans. Until then, the current 50% increase will remain in place.”

At first glance, a 25% permanent increase sounds like good news. But that 25% is measured against Claude Code’s original baseline — not against the temporary 50% boost users have been enjoying since May. Once the promotional period ends, the effective allowance drops.

The company added that “Compared to today, this works out to a 17% reduction in weekly limits on Claude Code. We’re working on exciting changes that will make it feel like you’re getting more from Claude, while having more visibility and control of your usage. Can’t wait to share them.”

Users and commentators quickly called out the mismatch between the celebratory framing and the actual outcome, with many noting it’s a familiar pattern as AI providers manage rising compute costs while trying to keep subscribers happy.

Nvidia to acquire Hugging Face at approximately $12.9 billion

Nvidia has agreed to acquire Hugging Face in a deal valued at approximately $12.9 billion, marking one of the largest acquisitions in the chipmaker’s history. The move expands Nvidia’s dominant footprint beyond AI hardware into the critical middleware and open-source developer layer.

Often described as the “GitHub of AI,” Hugging Face hosts over three million public models and a vast ecosystem of datasets. By controlling the primary distribution platform for open-weight models, Nvidia tightens the integration between its hardware architectures (such as CUDA and DGX Cloud) and the everyday workflows of millions of AI engineers.

The acquisition serves two major strategic purposes: it insulates Nvidia against closed-source competitors building proprietary silicon and establishes a direct software pipeline to commercial deployment. While the premium price tag—nearly triple Hugging Face’s 2023 valuation—reflects intense competition for developer mindshare, the deal also brings heightened scrutiny regarding platform neutrality and open-source governance.

Ultimately, absorbing Hugging Face cements Nvidia’s evolution from a hardware vendor into the end-to-end infrastructure backbone of artificial intelligence.

Source: TechCrunch

Google Launches Gemini 3.5 Transcribe for Precise Multilingual Speech-to-Text

From blog.google;

Today, we’re introducing Gemini 3.5 Transcribe, our most precise speech-to-text model yet, designed for intelligent voice interactions. Unlike conventional speech recognition models that struggle with background noise, complex jargon, and disfluency cleanup, Gemini 3.5 Transcribe converts raw audio directly into accurate, polished, formatted text.

Key Technical Capabilities

  • Automated Code-Switching & Language Detection: Supports over 85 languages and regional locales automatically. It dynamically handles sentence-level and intra-sentence language switching (code-mixing) without requiring pre-configured language hints.
  • Smart Transcription & Disfluency Removal: Automatically strips out filler words (e.g., “um,” “ah”), cleans up mid-sentence self-corrections, and normalizes spoken text into structured formats—converting phrases like “twenty five dollars” directly into “$25”.
  • Custom Vocabulary Biasing: Developers can supply up to 1,000 custom terms, acronyms, or domain-specific jargon phrases to ensure accurate recognition of specialized names and terminology.
  • Speed & Accuracy Benchmarks: Delivers a 70% faster time-to-final-transcription compared to Chirp 3. According to Artificial Analysis measurements, it achieves a 2.6% Word Error Rate (WER) on non-streaming tasks and a 4.0% WER in streaming mode.
  • Speaker Diarization & Word-Level Timestamps: Attributes speech segments for up to 8 distinct speakers and generates start and end time offsets down to the individual word for precise media indexing.

Dual Integration Paths

Feature / Model gemini-3.5-transcribe (Interactions API) gemini-3.5-transcribe-live (Live API)
Primary Use Case Recorded audio files, meeting logs, interviews Sub-second real-time streaming, voice agents
Max Audio Length Up to 1 hour (30 min with timestamps/diarization) 10 minutes per active session
Word-Level Timestamps Supported Not Supported
Speaker Diarization Supported (Up to 8 speakers) Not Supported

Consumer & Ecosystem Availability

In addition to developer access via Google AI Studio and Google Antigravity, Google is integrating Gemini 3.5 Transcribe across its wider product ecosystem:

  • Android (Gboard Rambler): Powers “Rambler,” converting long-form spoken thoughts into formatted notes and allowing voice-driven style updates.
  • Desktop & Browsers: Drives voice interaction in the Gemini macOS app, integrates screen context via Antigravity to increase spelling accuracy, and will soon enable direct voice-to-text dictation across Chrome web fields.

Welcome to “the boring bot!”

There is an overwhelming amount of noise surrounding Artificial Intelligence right now. Every single day brings another wave of breathless headlines, fear-mongering sci-fi predictions, and revolutionary breakthroughs that often turn out to be nothing more than impressive demo videos.

That is exactly why I built “The Boring Bot.”

I wanted to create a dedicated space to strip away the sensationalism and document what actually matters: real learnings, hands-on resources, and straightforward news across the entire AI ecosystem.

Here, “boring” is a badge of honor. It means we focus on utility, stability, implementation, and factual intelligence over internet drama.

What You Can Expect Here
Whether you are a developer building production-ready apps, an IT leader evaluating cloud infrastructure, or simply an enthusiast trying to make sense of the rapid pace of change, this platform is set up to cover AI from every practical angle:

  • Hardware & Compute: Unpacking the silicon powering the shift—from enterprise data center GPUs and custom ASICs to local workstation chips, edge runtimes, and local inference specs.
  • Software & Frameworks: Hands-on notes, model benchmarks, open-source libraries, developer APIs, and agentic workflows that you can actually deploy today.
  • Cloud & Infrastructure: Straightforward breakdowns of model hosting, vector databases, server monitoring, and scaling pipelines without hidden cost traps.
  • The Companies & Market: No-BS updates on the major players, emerging startups, open-source maintainers, regulation policies, and industry shifts that impact real-world tech stack decisions.
  • Curated Training & Resources: A growing hub of practical documentation, implementation guides, and tutorials to help you sharpen your technical skills.

A Quick Mandatory Disclosure:

Let’s be honest—it wouldn’t truly be a site about AI if I didn’t actually use AI to help draft, structure, and polish the articles here. Rest assured, human hands are at the steering wheel, but the bots are definitely doing the heavy lifting in the engine room. Think of it as drinking our own Kool-Aid, purely in the name of efficiency.

My Goal
This site exists to serve as a reliable, noise-free notebook and news hub. If a piece of tech is useful, efficient, and solves a real problem, you will read about it here—without the hype.

Thank you for dropping in for post number one. Bookmark the page, subscribe to updates, and let’s get down to the practical work of building with AI.