<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0">
  <channel>
    <title>LLM Pulse</title>
    <link>https://models.sutraworks.ai</link>
    <description>Every AI model. Every provider. Every change. Changes to AI model prices, capabilities and availability, plus daily model news.</description>
    <lastBuildDate>Fri, 21 Aug 2026 14:51:47 GMT</lastBuildDate>
    <item>
      <title>Crash Course to Master DeepSeek Harness from Scratch, Take You to Tinker with Cyber Lego</title>
      <link>https://eu.36kr.com/en/p/3948691750059394</link>
      <guid isPermaLink="false">98n7ya</guid>
      <pubDate>Fri, 21 Aug 2026 05:00:00 GMT</pubDate>
      <description>For example, click the dialog box, select any folder on your local computer, which can be a newly created one or one containing specific project files; after selection, we can put forward requirements to DeepSeek Harness.

Generally speaking, you can select the DeepSeek-V4-Pro model and set the reasoning level to Max. Although it costs a little more, it can also reduce the time we spend negotiating with AI.</description>
    </item>
    <item>
      <title>Indian IT Sector Pivots To Performance Billing As AI Demands Price Cuts - Whalesbook</title>
      <link>https://www.whalesbook.com/news/English/technology/Indian-IT-Sector-Pivots-To-Performance-Billing-As-AI-Demands-Price-Cuts/6a87975384d2dd5c12de796b</link>
      <guid isPermaLink="false">1a0z389</guid>
      <pubDate>Fri, 21 Aug 2026 00:09:55 GMT</pubDate>
      <description># Indian IT Sector Pivots To Performance Billing As AI Demands Price Cuts. India's IT services sector is moving from hourly billing to performance-based contracts as AI reduces the need for human labor. Clients are demanding price cuts of up to 30%, which is putting pressure on profit margins. The Indian IT services industry is undergoing a structural change as artificial intelligence (AI) forces a departure from the traditional model of charging clients by the hour. They are instead demanding contracts tied to specific business outcomes, effectively forcing IT firms to do more work for less cost. Clients are increasingly asking for price reductions in the range of 25% to 30%. This means the IT firm's revenue is now linked to the value it provides, such as efficiency gains, rather than the number of hours its employees work. IT firms are facing revenue headwinds as legacy managed-services contracts shrink faster than new, AI-led projects can scale.</description>
    </item>
    <item>
      <title>provider_added: DeepSeek: DeepSeek V4 Flash Vision Exp via kilo</title>
      <link>https://models.sutraworks.ai/changelog</link>
      <guid isPermaLink="false">002041311f28</guid>
      <pubDate>Fri, 21 Aug 2026 00:00:00 GMT</pubDate>
      <description>provider_added: DeepSeek: DeepSeek V4 Flash Vision Exp via kilo</description>
    </item>
    <item>
      <title>provider_added: DeepSeek V4 Flash Vision Exp via vercel</title>
      <link>https://models.sutraworks.ai/m/deepseek/deepseek-v4-flash-vision-exp</link>
      <guid isPermaLink="false">0f092fad7fa1</guid>
      <pubDate>Fri, 21 Aug 2026 00:00:00 GMT</pubDate>
      <description>provider_added: DeepSeek V4 Flash Vision Exp via vercel</description>
    </item>
    <item>
      <title>repriced: Gemma 4 26B A4B via nano-gpt</title>
      <link>https://models.sutraworks.ai/m/google/gemma-4-26b-a4b-it</link>
      <guid isPermaLink="false">20950a3a88b4</guid>
      <pubDate>Fri, 21 Aug 2026 00:00:00 GMT</pubDate>
      <description>cost.input: 0.13 → 0.08; cost.output: 0.4 → 0.33; cost.cache_read: 0.065 → 0.04</description>
    </item>
    <item>
      <title>repriced: GLM-5 via hyper</title>
      <link>https://models.sutraworks.ai/changelog</link>
      <guid isPermaLink="false">320374ad4c80</guid>
      <pubDate>Fri, 21 Aug 2026 00:00:00 GMT</pubDate>
      <description>cost.input: 0.83 → 0.91; cost.output: 2.558 → 2.934; cost.cache_write: 0.415 → 0.455</description>
    </item>
    <item>
      <title>repriced: Kimi K2.5 via hyper</title>
      <link>https://models.sutraworks.ai/changelog</link>
      <guid isPermaLink="false">4de7af559983</guid>
      <pubDate>Fri, 21 Aug 2026 00:00:00 GMT</pubDate>
      <description>cost.input: 0.5504 → 0.5444; cost.output: 2.885 → 2.855; cost.cache_write: 0.2752 → 0.2722</description>
    </item>
    <item>
      <title>capability_changed: Ox Alpha Free (Unlimited) via opencode</title>
      <link>https://models.sutraworks.ai/changelog</link>
      <guid isPermaLink="false">5a0b5a4c862a</guid>
      <pubDate>Fri, 21 Aug 2026 00:00:00 GMT</pubDate>
      <description>name: "Ox Alpha Free" → "Ox Alpha Free (Unlimited)"</description>
    </item>
    <item>
      <title>provider_added: Ox Alpha Free (Unlimited) via opencode-go</title>
      <link>https://models.sutraworks.ai/changelog</link>
      <guid isPermaLink="false">5b2a872d4966</guid>
      <pubDate>Fri, 21 Aug 2026 00:00:00 GMT</pubDate>
      <description>provider_added: Ox Alpha Free (Unlimited) via opencode-go</description>
    </item>
    <item>
      <title>repriced: Gemini 3.6 Flash via ofox</title>
      <link>https://models.sutraworks.ai/m/google/gemini-3.6-flash</link>
      <guid isPermaLink="false">61a77a6a84ab</guid>
      <pubDate>Fri, 21 Aug 2026 00:00:00 GMT</pubDate>
      <description>cost.input: 1.5 → 0.75; cost.output: 7.5 → 3.75; cost.cache_read: 0.15 → 0.075; cost.cache_write: 0.083 → 0.0415</description>
    </item>
    <item>
      <title>repriced: GPT-5.6 Sol via github-copilot</title>
      <link>https://models.sutraworks.ai/changelog</link>
      <guid isPermaLink="false">6743d7e0c76c</guid>
      <pubDate>Fri, 21 Aug 2026 00:00:00 GMT</pubDate>
      <description>cost.input: 5 → 2.5; cost.output: 30 → 15; cost.cache_read: 0.5 → 0.25; cost.cache_write: 6.25 → 3.125</description>
    </item>
    <item>
      <title>repriced: DeepSeek V4 Flash via openrouter</title>
      <link>https://models.sutraworks.ai/m/deepseek/deepseek-v4-flash</link>
      <guid isPermaLink="false">6eff15dbe370</guid>
      <pubDate>Fri, 21 Aug 2026 00:00:00 GMT</pubDate>
      <description>cost.input: 0.0826 → 0.08106; cost.output: 0.1652 → 0.16212; cost.cache_read: 0.01652 → 0.016212</description>
    </item>
    <item>
      <title>context_changed: DeepSeek V4 Flash 0731 via kilo</title>
      <link>https://models.sutraworks.ai/m/deepseek/deepseek-v4-flash-0731</link>
      <guid isPermaLink="false">7763dd2e80b3</guid>
      <pubDate>Fri, 21 Aug 2026 00:00:00 GMT</pubDate>
      <description>limit.output: 393216 → 384000</description>
    </item>
    <item>
      <title>provider_added: DeepSeek V4 Flash Vision Exp via opencode-go</title>
      <link>https://models.sutraworks.ai/changelog</link>
      <guid isPermaLink="false">790075e47011</guid>
      <pubDate>Fri, 21 Aug 2026 00:00:00 GMT</pubDate>
      <description>provider_added: DeepSeek V4 Flash Vision Exp via opencode-go</description>
    </item>
    <item>
      <title>provider_added: DeepSeek V4 Flash Vision Exp via edenai</title>
      <link>https://models.sutraworks.ai/m/deepseek/deepseek-v4-flash-vision-exp</link>
      <guid isPermaLink="false">794eeaab4750</guid>
      <pubDate>Fri, 21 Aug 2026 00:00:00 GMT</pubDate>
      <description>provider_added: DeepSeek V4 Flash Vision Exp via edenai</description>
    </item>
    <item>
      <title>provider_added: Qwen 3.8 27B Uncensored via nano-gpt</title>
      <link>https://models.sutraworks.ai/changelog</link>
      <guid isPermaLink="false">7f9d4e00d34d</guid>
      <pubDate>Fri, 21 Aug 2026 00:00:00 GMT</pubDate>
      <description>provider_added: Qwen 3.8 27B Uncensored via nano-gpt</description>
    </item>
    <item>
      <title>provider_added: Qwen3.8 27B via ofox</title>
      <link>https://models.sutraworks.ai/changelog</link>
      <guid isPermaLink="false">80f2d437f62d</guid>
      <pubDate>Fri, 21 Aug 2026 00:00:00 GMT</pubDate>
      <description>provider_added: Qwen3.8 27B via ofox</description>
    </item>
    <item>
      <title>repriced: Gemma 4 26B A4B IT via hyper</title>
      <link>https://models.sutraworks.ai/changelog</link>
      <guid isPermaLink="false">93edaffa8b08</guid>
      <pubDate>Fri, 21 Aug 2026 00:00:00 GMT</pubDate>
      <description>cost.input: 0.12 → 0.11; cost.output: 0.42 → 0.408; cost.cache_write: 0.06 → 0.055</description>
    </item>
    <item>
      <title>context_changed: Nemotron 3 Super 120B A12B via edenai</title>
      <link>https://models.sutraworks.ai/changelog</link>
      <guid isPermaLink="false">9bab89ea0fbe</guid>
      <pubDate>Fri, 21 Aug 2026 00:00:00 GMT</pubDate>
      <description>limit.context: 8000 → 262144</description>
    </item>
    <item>
      <title>repriced: DeepSeek V4 Flash via edenai</title>
      <link>https://models.sutraworks.ai/m/deepseek/deepseek-v4-flash</link>
      <guid isPermaLink="false">a06f9408bc8d</guid>
      <pubDate>Fri, 21 Aug 2026 00:00:00 GMT</pubDate>
      <description>cost.input: 0.44 → 0.22; cost.output: 1.32 → 0.66; cost.cache_read: 0.014 → 0.007</description>
    </item>
    <item>
      <title>repriced: Kimi K2.6 via openrouter</title>
      <link>https://models.sutraworks.ai/m/moonshotai/kimi-k2.6</link>
      <guid isPermaLink="false">a23ff06910f7</guid>
      <pubDate>Fri, 21 Aug 2026 00:00:00 GMT</pubDate>
      <description>cost.input: 0.95 → 0.5795; cost.output: 4 → 2.44; cost.cache_read: 0.16 → 0.0976</description>
    </item>
    <item>
      <title>capability_changed: DeepSeek V4 Flash Vision Exp via kilo</title>
      <link>https://models.sutraworks.ai/m/deepseek/deepseek-v4-flash-vision-exp</link>
      <guid isPermaLink="false">c9627deeb41c</guid>
      <pubDate>Fri, 21 Aug 2026 00:00:00 GMT</pubDate>
      <description>name: "DeepSeek: DeepSeek V4 Flash Vision Exp" → "DeepSeek V4 Flash Vision Exp"; family: "deepseek" → "deepseek-flash"</description>
    </item>
    <item>
      <title>repriced: DeepSeek V4 Pro via edenai</title>
      <link>https://models.sutraworks.ai/m/deepseek/deepseek-v4-pro</link>
      <guid isPermaLink="false">d0e6d6958eee</guid>
      <pubDate>Fri, 21 Aug 2026 00:00:00 GMT</pubDate>
      <description>cost.input: 1.32 → 0.66; cost.output: 3.96 → 1.98; cost.cache_read: 0.044 → 0.022</description>
    </item>
    <item>
      <title>repriced: Gemma 4 31B via nano-gpt</title>
      <link>https://models.sutraworks.ai/m/google/gemma-4-31b-it</link>
      <guid isPermaLink="false">d9b1ea7dc0f0</guid>
      <pubDate>Fri, 21 Aug 2026 00:00:00 GMT</pubDate>
      <description>cost.input: 0.1 → 0.08; cost.output: 0.35 → 0.33; cost.cache_read: 0.05 → 0.04</description>
    </item>
    <item>
      <title>provider_added: DeepSeek V4 Flash (RanoAI) via llmgateway-providers</title>
      <link>https://models.sutraworks.ai/changelog</link>
      <guid isPermaLink="false">daff5c3ac49f</guid>
      <pubDate>Fri, 21 Aug 2026 00:00:00 GMT</pubDate>
      <description>provider_added: DeepSeek V4 Flash (RanoAI) via llmgateway-providers</description>
    </item>
    <item>
      <title>repriced: DeepSeek V4 Flash 0731 via openrouter</title>
      <link>https://models.sutraworks.ai/m/deepseek/deepseek-v4-flash-0731</link>
      <guid isPermaLink="false">e0e585e9609a</guid>
      <pubDate>Fri, 21 Aug 2026 00:00:00 GMT</pubDate>
      <description>limit.output: 393216 → 384000; cost.input: 0.14 → 0.08; cost.output: 0.28 → 0.18; cost.cache_read: 0.028 → 0.016</description>
    </item>
    <item>
      <title>capability_changed: DeepSeek V4 Flash Vision Exp via openrouter</title>
      <link>https://models.sutraworks.ai/m/deepseek/deepseek-v4-flash-vision-exp</link>
      <guid isPermaLink="false">e6f21ba726f0</guid>
      <pubDate>Fri, 21 Aug 2026 00:00:00 GMT</pubDate>
      <description>family: "deepseek" → "deepseek-flash"</description>
    </item>
    <item>
      <title>repriced: Gemini 3.7 Flash via ofox</title>
      <link>https://models.sutraworks.ai/m/google/gemini-3.7-flash</link>
      <guid isPermaLink="false">eaa3bedd67ad</guid>
      <pubDate>Fri, 21 Aug 2026 00:00:00 GMT</pubDate>
      <description>cost.input: 1.5 → 0.75; cost.output: 7.5 → 3.75; cost.cache_read: 0.15 → 0.075; cost.cache_write: 0.083 → 0.0415</description>
    </item>
    <item>
      <title>model_added: DeepSeek V4 Flash Vision Exp via openrouter</title>
      <link>https://models.sutraworks.ai/m/deepseek/deepseek-v4-flash-vision-exp</link>
      <guid isPermaLink="false">f0163350a8af</guid>
      <pubDate>Fri, 21 Aug 2026 00:00:00 GMT</pubDate>
      <description>model_added: DeepSeek V4 Flash Vision Exp via openrouter</description>
    </item>
    <item>
      <title>provider_added: Qwen 3.6 35B A3B Uncensored via nano-gpt</title>
      <link>https://models.sutraworks.ai/changelog</link>
      <guid isPermaLink="false">f723f96ecc1e</guid>
      <pubDate>Fri, 21 Aug 2026 00:00:00 GMT</pubDate>
      <description>provider_added: Qwen 3.6 35B A3B Uncensored via nano-gpt</description>
    </item>
    <item>
      <title>provider_added: Meituan: LongCat 2.0 (free) via kilo</title>
      <link>https://models.sutraworks.ai/changelog</link>
      <guid isPermaLink="false">ff97ae8ec1c1</guid>
      <pubDate>Fri, 21 Aug 2026 00:00:00 GMT</pubDate>
      <description>provider_added: Meituan: LongCat 2.0 (free) via kilo</description>
    </item>
    <item>
      <title>DeepSeek V4 Pro vs Qwen3.8 vs Glimmer: 80x Params [2026]</title>
      <link>https://tech-insider.org/ca/deepseek-v4-pro-vs-qwen3-8-max-vs-muse-glimmer-2026</link>
      <guid isPermaLink="false">ubiwlq</guid>
      <pubDate>Fri, 21 Aug 2026 00:00:00 GMT</pubDate>
      <description>DeepSeek shipped the general-availability “0813” build of V4 Pro on August 13, 2026, replacing the April preview that had been running behind the same model ID. It is a mixture-of-experts model with roughly 1.6 trillion total parameters and about 49 billion active per token, built on a DeepSeekV4ForCausalLM architecture class, according to the model card DeepSeek published on Hugging Face. The headline engineering story is the attention mechanism: DeepSeek combined Compressed Sparse Attention [...] DeepSeek V4 Pro 0813 is available two ways: as hosted API access through DeepSeek’s own platform using OpenAI- or Anthropic-compatible request formats, or as a direct download from Hugging Face under the model ID `deepseek-ai/DeepSeek-V4-Pro-0813` for teams with the GPU capacity to self-host. Several router services, including OpenRouter, also proxy the hosted API for teams that want a single integration point across multiple model providers. [...] cache-hit price is close to free. Modality-wise, DeepSeek’s own spec pages describe V4 Pro 0813 as text-only, though at least one third-party comparison mentions experimental image reasoning in some deployments; treat vision support as unconfirmed until DeepSeek documents it directly. Weights are open under the MIT license and hosted on Hugging Face at `deepseek-ai/DeepSeek-V4-Pro-0813`, making it fully self-hostable for commercial use.</description>
    </item>
    <item>
      <title>Grok 4.6 Matches the Frontier Models at a 60% Discount. Now SpaceX Has the Developers to Use It. - The Globe and Mail</title>
      <link>https://www.theglobeandmail.com/investing/markets/stocks/GOOG/pressreleases/3958396/grok-46-matches-the-frontier-models-at-a-60-discount-now-spacex-has-the-developers-to-use-it</link>
      <guid isPermaLink="false">p5gfo5</guid>
      <pubDate>Thu, 20 Aug 2026 20:00:00 GMT</pubDate>
      <description>## Key Points

 SpaceX's purchase of Cursor marks its arrival as a serious contender in the software development arena.
 Grok 4.6 is the first xAI model to rank alongside the best from OpenAI and Anthropic on key benchmarks.
 SpaceX's cost advantage comes from owning its computing power rather than renting it.
 10 stocks we like better than Space Exploration Technologies › [...] SpaceX(NASDAQ: SPCX) closed its $60 billion acquisition of Cursor last week, just days after releasing its latest AI model, Grok 4.6. The deal gives SpaceX an instant foothold in the enterprise AI market through Cursor's popular code editor. [...] An image of Elon Musk in the White House.

SpaceX CEO Elon Musk. Image source: The Motley Fool.

## Grok is in the game

Grok 4.6 scored 61 on the benchmark Artificial Analysis Intelligence Index, matching GPT-5.6 Sol and sitting just behind the newest Claude models. On the SWE-bench that ranks large language models as tools for coding, it's in the top cluster.</description>
    </item>
    <item>
      <title>Grok 4.6 Matches the Frontier Models at a 60% Discount. Now SpaceX Has the Developers to Use It.</title>
      <link>https://www.fool.com/investing/2026/08/20/grok-46-matches-the-frontier-models-at-a-60-discou</link>
      <guid isPermaLink="false">1f03p6o</guid>
      <pubDate>Thu, 20 Aug 2026 19:53:00 GMT</pubDate>
      <description>## Key Points

 SpaceX's purchase of Cursor marks its arrival as a serious contender in the software development arena.
 Grok 4.6 is the first xAI model to rank alongside the best from OpenAI and Anthropic on key benchmarks.
 SpaceX's cost advantage comes from owning its computing power rather than renting it. [...] Grok 4.6 scored 61 on the benchmark Artificial Analysis Intelligence Index, matching GPT-5.6 Sol and sitting just behind the newest Claude models. On the SWE-bench that ranks large language models as tools for coding, it's in the top cluster. [...] Accessibility Menu

▲ S&amp;P 500 +---%|▲ Stock Advisor +---%Join The Motley Fool

AccessibilityHelp

The Motley FoolThe Motley Fool

SpaceX (SPCX -4.05%) closed its $60 billion acquisition of Cursor last week, just days after releasing its latest AI model, Grok 4.6. The deal gives SpaceX an instant foothold in the enterprise AI market through Cursor's popular code editor.</description>
    </item>
    <item>
      <title>Grok 4.6 Is Here: Video Explanations and Agentic AI, Explained</title>
      <link>https://www.basenor.com/blogs/news/grok-4-6-is-here-video-explanations-and-agentic-ai-explained</link>
      <guid isPermaLink="false">oysopn</guid>
      <pubDate>Thu, 20 Aug 2026 18:07:06 GMT</pubDate>
      <description>Grok 4.6 is the latest version of xAI's Grok AI model, officially released on August 12, 2026. Musk flagged it publicly on August 20. According to xAI, the update focuses on three areas: upgraded multimodal understanding (text, images, and video), long-running agentic work, and the introduction of Grok Bot — a persistent AI agent that runs on a cloud computer rather than in a single session window.

### What does Grok Bot actually do? [...] According to Musk, image and video understanding in Grok 4.6 has undergone significant upgrades. The model can now handle visual inputs alongside natural language with greater accuracy — meaning you can share an image or video clip and expect more precise, contextually aware analysis in return. For practical use cases, this matters for anything from reading charts and diagrams to interpreting real-world footage.

### Is this relevant to Tesla owners specifically? [...] Elon Musk announced Grok 4.6 on August 20, pointing to a significant capability jump for xAI's conversational AI — including the ability to generate videos that explain complex concepts. The update, which officially launched on August 12 according to xAI, pushes Grok further into multimodal and agentic territory, moving it well beyond a simple chat interface.

Elon Musk announces Grok 4.6 on X

### What is Grok 4.6, and when did it launch?</description>
    </item>
    <item>
      <title>Ramp launches its own AI model router, called Router - TechCrunch</title>
      <link>https://techcrunch.com/2026/08/20/ramp-launches-its-own-ai-model-router-called-router/</link>
      <guid isPermaLink="false">ij3lkq</guid>
      <pubDate>Thu, 20 Aug 2026 16:46:00 GMT</pubDate>
      <description>Corporate expense management platform Ramp is hot on the heels of Stripe in setting up toll houses for AI inference. Ramp on Wednesday evening launched its own AI model routing service, dubbed Router, that lets users and companies use and switch between various large language models through an API. It’s free to use for the remainder of 2026 (users will still have to pay for AI model inference costs), and it comes with a $26 credit launch offer. In its function, Router is pretty similar to how OpenRouter operates, though the latter offers many more AI model options than Ramp’s current offerings. For Ramp, entering the model routing business offers a two-pronged opportunity: it gets to tap the booming AI inference market, and *also* offer its existing clients a model routing service that fits in neatly with its existing products, which includes AI token usage monitoring and token spend management.</description>
    </item>
    <item>
      <title>GPT-5.6 vs DeepSeek V4 Pro 0813: 714x Cheaper Input [2026]</title>
      <link>https://tech-insider.org/gpt-5-6-vs-deepseek-v4-pro-0813-2026</link>
      <guid isPermaLink="false">1wtisdj</guid>
      <pubDate>Thu, 20 Aug 2026 12:10:05 GMT</pubDate>
      <description>DeepSeek V4 Pro 0813, by contrast, is a single open-weight model, a refresh of the earlier DeepSeek V4 Pro release, dated to its August 13 build number. It is DeepSeek’s flagship reasoning and coding model, sitting above the cheaper V4 Flash tier the same way GPT-5.6 Sol sits above Luna. Where GPT-5.6 is API-only and fully closed, DeepSeek V4 Pro 0813 ships weights that developers can download, self-host, or run through third-party inference providers like OpenRouter, which changes the cost [...] DeepSeek’s strongest published reasoning gains, by contrast, cluster around legal and formal-proof tasks, plus general knowledge breadth. The Vals Index puts DeepSeek V4 Pro 0813’s overall score at 52.37%, a jump of 9.48 points over the prior DeepSeek V4 baseline’s 42.89%, moving it to #18 overall on that leaderboard. On the narrower BenchLM DeepSeek-only leaderboard, updated August 19, 2026, V4 Pro 0813 is ranked the top DeepSeek model with a score of 61.3, ahead of every earlier DeepSeek [...] Both models land in roughly the same context window class. GPT-5.6’s three tiers all carry approximately 1.05 million tokens of context, according to OpenAI’s model documentation pages. DeepSeek V4 Pro 0813’s context window is listed at 1.0 million tokens in the most recent cost calculator explicitly tied to the 0813 build, though it is worth noting that some older DeepSeek pricing guides still reference a smaller 256K window for a prior V4 Pro release, a discrepancy that likely reflects the</description>
    </item>
    <item>
      <title>How to Set Up DeepSeek V4 Pro: 12 Steps, 90 Min [2026]</title>
      <link>https://tech-insider.org/how-to-set-up-deepseek-v4-pro-2026</link>
      <guid isPermaLink="false">3zgsoh</guid>
      <pubDate>Thu, 20 Aug 2026 10:09:36 GMT</pubDate>
      <description>DeepSeek positions V4 Pro as its agentic and coding-first model, aimed squarely at complex reasoning chains, tool-calling workflows, and long-document analysis. The cheaper sibling, DeepSeek V4 Flash (`deepseek-v4-flash`, updated to the 0731 build on July 31, 2026), trades some reasoning depth for a 5x higher concurrency ceiling and a fraction of the per-token cost. We already covered Flash setup in a dedicated DeepSeek V4-Flash tutorial; this guide focuses on the Pro tier and the extra steps [...] DeepSeek V4 Pro’s positioning as an agentic model means tool calling (also called function calling) is where it’s meant to earn its higher price over Flash. The API follows the same tool-calling schema as OpenAI’s Chat Completions format, so if you’ve built agents against GPT models before, the pattern will look familiar. Here’s a minimal agent that gives the model access to a weather-lookup function and lets it decide when to call it:</description>
    </item>
    <item>
      <title>GPT-5.6 vs Gemini 3.7 Flash vs Grok 4.6: 7x Price Gap [2026]</title>
      <link>https://tech-insider.org/ca/gpt-5-6-vs-gemini-3-7-flash-vs-grok-4-6-2026</link>
      <guid isPermaLink="false">1e0bws8</guid>
      <pubDate>Thu, 20 Aug 2026 10:00:00 GMT</pubDate>
      <description>xAI’s Grok 4.6 continues a faster release cadence than either rival, with Grok 4.5 having entered beta only months earlier at SpaceX and Tesla before Grok 4.6’s general release. That speed comes with a tradeoff: xAI’s benchmark documentation for Grok 4.6 is the least consistent of the three across independent trackers, and its context window still trails the roughly one-million-token standard that OpenAI and Google have both settled on for their current flagship-adjacent tiers. [...] xAI API curl  \ -H "Authorization: Bearer $XAI_API_KEY" \ -H "Content-Type: application/json" \ -d '{ "model": "grok-4.6", "messages": [{"role": "user", "content": "Summarize this contract."}] }'</description>
    </item>
    <item>
      <title>Google Goes Back to School With New AI Study Tools - cnet.com</title>
      <link>https://www.cnet.com/tech/google-gemini-search-ai-study-tools/</link>
      <guid isPermaLink="false">1jblvrz</guid>
      <pubDate>Thu, 20 Aug 2026 00:06:51 GMT</pubDate>
      <description>A new dedicated study hub includes a study notebook, customized flash cards and practice quizzes. It’s back-to-school season, and Google is getting in on the academic action. The company announced on Wednesday that a dedicated student hub, featuring a suite of AI study tools, will now be available on its Gemini AI assistant and through Google Search. The dedicated study hub can be found within Gemini and contains a suite of tools, including a study notebook, customized flash cards and practice quizzes to help with research and assignments. Students will be able to ask Gemini to dig into a topic they’re researching and then switch tasks while the tool compiles the information in the background. When the report is finished, Gemini will send a notification and be available to discuss the report’s details and participate in a question-and-answer session to help the student better understand the findings. Stick with AI Mode in Google Search, and you can ask follow-up questions about the topic in question to dig deeper into a subject.</description>
    </item>
    <item>
      <title>Amazon Adds Grok 4.6 to Bedrock</title>
      <link>https://www.tradingview.com/news/gurufocus:6994020ce094b:0-amazon-adds-grok-4-6-to-bedrock</link>
      <guid isPermaLink="false">fr46mc</guid>
      <pubDate>Wed, 19 Aug 2026 21:58:24 GMT</pubDate>
      <description>GuruFocus
GuruFocus

# Amazon Adds Grok 4.6 to Bedrock

Amazon.com Inc. (AMZN, Financials), the cloud and e-commerce giant, is adding SpaceXAI's latest Grok model to Bedrock as competition for enterprise AI customers continues to heat up.

Grok 4.6 is now generally available to developers in supported AWS regions, SpaceXAI said. This opens up a much wider path into companies that are already building AI applications on Amazon Web Services. [...] Amazon Bedrock enables customers to access and deploy third-party AI models without having to build the underlying infrastructure themselves. Adding Grok provides AWS users with another option, on top of the growing list of models already available from the platform.

But for SpaceXAI, the bigger win is dissemination. AWS could be an opportunity to get Grok in front of enterprise developers and turn model improvements into real commercial usage, not just consumer attention.</description>
    </item>
    <item>
      <title>AI in Medicine: Study Weighs Risks, Safe Use Strategies - Mirage News</title>
      <link>https://www.miragenews.com/ai-in-medicine-study-weighs-risks-safe-use-1730018/</link>
      <guid isPermaLink="false">1iv1bsm</guid>
      <pubDate>Wed, 19 Aug 2026 20:30:00 GMT</pubDate>
      <description>Large language models (LLMs), including the models behind ChatGPT and Claude as well as numerous other systems developed specifically for medical applications, are increasingly used in clinical workflows. However, their adoption is outpacing the development of systems for oversight and safety. An interdisciplinary team of researchers at the Else Kröner Fresenius Center (EKFZ) for Digital Health at TUD Dresden University of Technology and University Hospital Dresden, together with national and international colleagues, has systematically analyzed the risks associated with LLM use in medicine. The review, published in Nature, brings together evidence from medical AI, cybersecurity, regulatory science, ethics and behavioral psychology and outlines strategies for trustworthy and responsible use of artificial intelligence (AI) in clinical practice. LLMs have the potential to support and enhance the work of healthcare professionals in a variety of areas. The authors of the newly published review show that these risks can arise throughout the entire lifecycle of AI systems: from initial model design to training data, model deployment, and real-world use in clinical environments.</description>
    </item>
    <item>
      <title>Google launches new study tools for students across Search and Gemini - TechCrunch</title>
      <link>https://techcrunch.com/2026/08/19/google-launches-new-study-tools-for-students-across-search-and-gemini/</link>
      <guid isPermaLink="false">1n3gxdx</guid>
      <pubDate>Wed, 19 Aug 2026 19:00:00 GMT</pubDate>
      <description>Google on Wednesday announced a slew of new study tools across Search and Gemini, including AI-generated interactive visuals, 3D simulations, a decided student hub, customized practice quizzes, and more. The launch of the new study features marks Google’s latest effort to make Gemini the AI assistant that students turn to when learning and studying, as it competes with companies like OpenAI and education startups such as Knowt and Gauth, which are also offering their own learning and practice tools. On Search, students can now generate custom tools and simulations to help them understand complex topics. For example, if a student is learning about pH levels, they can search for “pH scale” and get an interactive visual in an AI Overview that makes the basics easier to understand. To go even further, they can ask a follow-up question for something more specialized, like plotting citrus fruits on the pH scale, and AI Mode in Search will then create an interactive experience tailored to the question.</description>
    </item>
    <item>
      <title>Gemini Live adds Deep Research as Notebooks come to AI Mode - 9to5Google</title>
      <link>https://9to5google.com/2026/08/19/gemini-app-ai-mode-study-tools/</link>
      <guid isPermaLink="false">1p89hd8</guid>
      <pubDate>Wed, 19 Aug 2026 19:00:00 GMT</pubDate>
      <description># Gemini Live adds Deep Research as Notebooks come to AI Mode. With the back-to-school season underway, Google is introducing new study tools across the Gemini app and Search. College students in the US can get one free year of Google AI Pro ($19.99 per month), while there’s a bundle with YouTube Premium that offers up to 70% off. A new “Student” hub (gemini.google.com/students) in the Gemini side panel surfaces offers and tools, like flashcards and practice quizzes. Meanwhile, Google AI Overviews and AI Mode are adding a host of study tools. * “To go even further, you can follow up and ask to see something more specialized, like plotting citrus fruits on the pH scale, and AI Mode will create a customized experience just for your question.”. AI Mode is getting integration with Gemini Notebooks. Google Lens is getting a new interactive learning experience that can discuss “tough concepts and confirm if you’re on the right track.” Available in the coming weeks (globally in English), you just snap a photo of the problem.</description>
    </item>
    <item>
      <title>DeepSeek Open-Sources the Missing Layer Between AI Models and Agents</title>
      <link>https://www.hpcwire.com/aiwire/2026/08/19/deepseek-open-sources-the-missing-layer-between-ai-models-and-agents</link>
      <guid isPermaLink="false">5okk9b</guid>
      <pubDate>Wed, 19 Aug 2026 16:47:46 GMT</pubDate>
      <description>model. A sort of collection of software that determines whether an AI agent can actually get a job done. There are some hard numbers behind the release, too. Harness arrived alongside DeepSeek V4-Pro-0813, the latest version of the company’s flagship model, with much of the attention this time going to its agent capabilities. DeepSeek puts V4-Pro-0813 at 87.9 on Terminal Bench 2.1. It scored 74.1 on Toolathlon-Verified, 71.1 on DSBench-FullStack, and 67.2 on DSBench-Hard. These are DeepSeek’s [...] model. A sort of collection of software that determines whether an AI agent can actually get a job done. There are some hard numbers behind the release, too. Harness arrived alongside DeepSeek V4-Pro-0813, the latest version of the company’s flagship model, with much of the attention this time going to its agent capabilities. DeepSeek puts V4-Pro-0813 at 87.9 on Terminal Bench 2.1. It scored 74.1 on Toolathlon-Verified, 71.1 on DSBench-FullStack, and 67.2 on DSBench-Hard. These are DeepSeek’s [...] There are some hard numbers behind the release, too. Harness arrived alongside DeepSeek V4-Pro-0813, the latest version of the company’s flagship model, with much of the attention this time going to its agent capabilities. DeepSeek puts V4-Pro-0813 at 87.9 on Terminal Bench 2.1. It scored 74.1 on Toolathlon-Verified, 71.1 on DSBench-FullStack, and 67.2 on DSBench-Hard.</description>
    </item>
    <item>
      <title>DeepSeek V4 Flash has been criticized for its high benchmark scores but struggles in real-world tasks, particularly with orchestration. - GIGAZINE</title>
      <link>https://gigazine.net/gsc_news/en/20260819-deepseek-v4-flash-agent</link>
      <guid isPermaLink="false">z0gfa6</guid>
      <pubDate>Wed, 19 Aug 2026 10:00:00 GMT</pubDate>
      <description>While DeepSeek-V4 Flash achieved high scores in agent-based benchmarks, its success rate was lower on more challenging tasks. Taryn Prambu, a writer specializing in AI and cybersecurity, points out, 'This gap demonstrates why, in enterprise environments, the success or failure of a model depends not on the raw capabilities of the model itself, but on the 'orchestration' that brings multiple systems and tasks together. Even with the same model, the results can vary greatly depending on the [...] Technology analyst Karmi Levy stated that even after the price increase, DeepSeek remains 'far cheaper in every aspect' compared to competing models such as OpenAI, Anthropic, Google, and xAI, and emphasized the importance of appropriately allocating agent AI capabilities. For example, tasks like batch processing are often routine and repetitive, rather than requiring advanced technology, making it appropriate to use a cheaper and more efficient model. Levy said, 'Companies can consider [...] Sanchit Vil Gogia of Greyhound Research, a global technology research and consulting firm, said, 'What's important is whether the AI model's performance is sufficient to meet the actual workflows that companies operate, not whether it's the best at every benchmark. Companies need to determine which combination of model, harness, and provider will get the job done most securely and at the lowest cost.'</description>
    </item>
    <item>
      <title>GPT-5.6: Frontier intelligence that scales with your ambition</title>
      <link>https://openai.com/index/gpt-5-6</link>
      <guid isPermaLink="false">bxqevo</guid>
      <pubDate>Tue, 18 Aug 2026 22:00:00 GMT</pubDate>
      <description>GPT‑5.6 is our strongest model yet for accelerating AI research. Inside OpenAI, researchers use it across the development loop: diagnosing failures, optimizing training systems, running experiments, and interpreting results. We already saw that acceleration and stronger adoption during the internal testing period of GPT‑5.6, as average daily output tokens per active researcher were more than twice the highest level observed for GPT‑5.5.</description>
    </item>
    <item>
      <title>Alibaba’s lightweight Qwen takes on OpenAI, DeepSeek, Zhipu’s larger AI systems</title>
      <link>https://www.scmp.com/tech/tech-trends/article/3364404/alibabas-lightweight-qwen-model-takes-larger-ai-systems-openai-deepseek-zhipu</link>
      <guid isPermaLink="false">1cmsujr</guid>
      <pubDate>Tue, 18 Aug 2026 12:00:05 GMT</pubDate>
      <description>Xinmei Shen

Alibaba Group Holding’s new lightweight AI model Qwen3.8-27B has matched much larger near-frontier rivals while being able to run on everyday hardware, impressing developers as local AI gains momentum.

The Qwen3.8-27B, a small model with 27 billion parameters, performed on par with OpenAI’s GPT-5.6 Luna, which was billed as the most cost-efficient model in the US lab’s latest flagship series, benchmark firm Artificial Analysis said on Monday. [...] Artificial intelligence

TechTech Trends

# Alibaba’s lightweight Qwen model takes on larger AI systems from OpenAI, DeepSeek, Zhipu

Chinese tech giant’s latest offering performed on par with OpenAI’s GPT-5.6 Luna and nearly matched DeepSeek, Zhipu’s open-weight models

2-MIN READ2-MIN

Chinese tech titan Alibaba Group Holding’s new lightweight AI model Qwen3.8-27B has matched much larger rivals including OpenAI, DeepSeek and Zhipu. Photo: Shutterstock

Xinmei Shen [...] The new findings came days after Alibaba released Qwen3.8-27B’s model weights – the underlying parameters that encode its intelligence – last Friday.

On Artificial Analysis’ Agentic Index, which measures models’ performance in AI agent-focused workflows, Alibaba’s small model outperformed GPT-5.6 series’ mid-tier model Terra and Anthropic’s powerful Claude Opus 4.8 released in May.

Advertisement

Select Voice

Select Speed</description>
    </item>
    <item>
      <title>BABA Stock Gains On Launching QwenAI Model To Challenge Meta's Lead In Open-Source AI — Noticias de TradingView</title>
      <link>https://es.tradingview.com/news/stocktwits:0e3f9df13094b:0-baba-stock-gains-on-launching-qwenai-model-to-challenge-meta-s-lead-in-open-source-ai</link>
      <guid isPermaLink="false">4gwjuo</guid>
      <pubDate>Tue, 18 Aug 2026 10:00:00 GMT</pubDate>
      <description>Puntos clave:

 Alibaba introduced Qwen3.8-27B, a lightweight artificial intelligence model optimized to run on personal computers and local hardware.
 The Chinese tech giant publicly released the weights for its top-tier Qwen3.8 Max model, consolidating its lead in developer adoption over Western rivals.
 The double launch directly answers Meta’s recent release of its Muse Glimmer model family as both tech giants vie for control of the open-source developer ecosystem. [...] The e-commerce and cloud computing firm unveiled Qwen3.8-27B, saying the compact software delivers strong performance across coding, research, professional tasks, and complex agentic workflows. According to the company, the localized model matches systems ten times its size while running locally on personal devices rather than relying on distant data centers.</description>
    </item>
    <item>
      <title>Previewing GPT-5.6 Sol: a next-generation model | OpenAI</title>
      <link>https://openai.com/index/previewing-gpt-5-6-sol</link>
      <guid isPermaLink="false">sfmerg</guid>
      <pubDate>Tue, 18 Aug 2026 10:00:00 GMT</pubDate>
      <description>.5.5 6 5..6 6 5 5.5 5 5 6 5 6.6 6.5 6...6..5 6 5 5.5 6.6 6 5 6 [...] 5    .5 5 5     5  5             
           5                                      
         5   55     5  6  5.                      
                             5   6 5 55           
              5              5      5     6       
                5         6    5                  
       5                                 .  5     
           66              5           .   6 5    
      .  5             .  5    6    5 5      65 [...] 5 6 5 6 6 6 5 5 5 5 5 5 5 5 6 5..5.5 5 5 5 5 5 5..5 5..5 5 5 5 5 6 5 6 5 5 5 5.6 5 5.6 5 5 5 5 6 5 5.5 6 5 5.5 5 5 5 5 6 5 5 6 5 6.5 5 6 5 5 5 5.5 5 5 5 5.5 5 5 5 5 5 6 6 6 5 5.5 5 5 5 5.6 6 5 5 5 5 6 6 5 5 5.5 5 6 6 5 5 5 6 5 5 5 6 5 5 5 6 5 5 6 5 5 5.5 5.6 6 6 5 5 5 5 5 5 6 6 5.5 5 5 5</description>
    </item>
  </channel>
</rss>