Daily News

Daily AI news

A readable stream of AI updates for creator research. News stays separate from daily picks until it is verified and framed.

Jul 14, 2026Today

14
  • Jul 14, 2026

    official

    OpenAI: gpt-oss refresh

    OpenAI 7/13 gpt-oss refresh: 120B MoE + 20B dense post-training snapshot under Apache-2.0

    OpenAI index page on 2026-07-13 publishes a refresh of the gpt-oss open-weight line — 120B MoE and 20B dense variants with new post-training snapshots; license remains Apache-2.0 with the OpenAI usage-policy addendum. Self-reported benchmark numbers: MMLU-Pro 84.6, GPQA-Diamond 76.2, LiveCodeBench v6 72.1, SWE-Bench Verified 68.3, BFCL v3 65.4, Terminal-Bench 2.1 76.8. Semantic boundary: vendor index + Hugging Face community model card + a tech-press item (The Decoder) — not an independent reproduction of any specific score.

    1 daily topicOriginally published Jul 13, 2026View source
  • Jul 14, 2026

    official

    Anthropic: Claude Sonnet 4.5

    Anthropic 7/13 Claude Sonnet 4.5: mid-tier flagship for everyday tasks at the same $3/$15 pricing

    Anthropic newsroom on 2026-07-13 introduces Claude Sonnet 4.5 as the new mid-tier flagship positioned between Haiku 4.5 and Opus 4.5; 200K context, native tool-use with extended-thinking mode, $3 input / $15 output per million tokens (same as Sonnet 4.5 prior). Self-reported figures: SWE-Bench Verified 67.4, GPQA-Diamond 74.8, MMLU-Pro 82.9, BFCL v3 63.7, Terminal-Bench 2.1 74.2. Deployment: API + claude.ai + AWS Bedrock + Google Vertex AI. Semantic boundary: vendor newsroom + Hugging Face community card + a Decoder tech-press item — not an independent reproduction of any specific score.

    1 daily topicOriginally published Jul 13, 2026View source
  • Jul 14, 2026

    official

    Google: Gemini 3.5 Pro mid-July

    Google 7/13 Gemini 3.5 Pro mid-July: agent scaffolds + new image-edit-via-namespace endpoint

    Google blog on 2026-07-13 publishes a mid-July tier update for Gemini 3.5 Pro emphasizing agent scaffolds (server-side tool-use loop) + a new image-edit-via-namespace endpoint; pro-tier pricing unchanged. Self-reported figures: MMLU-Pro 86.3, GPQA-Diamond 79.1, LiveCodeBench v6 75.4, SWE-Bench Verified 71.2, BFCL v3 68.9, Terminal-Bench 2.1 78.5. Deployment: Gemini API + Vertex AI. Semantic boundary: vendor blog + Hugging Face community model card; no media corroboration today.

    1 daily topicOriginally published Jul 13, 2026View source
  • Jul 14, 2026

    official

    Meta: Llama 4 Behemoth refresh

    Meta 7/13 Llama 4 Behemoth refresh: 1M context + code-weight specialist under Llama 4 Community License

    Meta AI blog on 2026-07-13 publishes a refresh of the Llama 4 family emphasizing long-context (1M + token-efficiency) and a code-weight specialist. Self-reported figures: MMLU-Pro 83.4, GPQA-Diamond 75.6, LiveCodeBench v6 73.8, SWE-Bench Verified 69.7, HumanEval+ 91.2, Terminal-Bench 2.1 75.9. License: Llama 4 Community License with the commercial-use restrictions addendum. Semantic boundary: vendor blog + Hugging Face community card; benchmark numbers are vendor-supplied.

    1 daily topicOriginally published Jul 13, 2026View source
  • Jul 14, 2026

    official

    ByteDance Volcengine: Doubao 1.6 Pro

    ByteDance 7/13 Doubao 1.6 Pro: 256K context + tool-use refresh, full weights gated to Volcengine API

    ByteDance Volcengine product news on 2026-07-13 publishes a Doubao 1.6 Pro update — 256K context, refreshed tool-use schema, post-training emphasizing document-extraction / retrieval-grounded tasks. Self-reported figures: MMLU-Pro 80.7, GPQA-Diamond 72.4, C-Eval 88.6, CMMLU 86.9, LiveCodeBench v6 70.3, Terminal-Bench 2.1 73.5. License: commercial-only via Volcengine API; full weights remain gated. Semantic boundary: vendor product news + a Hugging Face mirror card listed as researcher preview only.

    1 daily topicOriginally published Jul 13, 2026View source
  • Jul 14, 2026

    official

    Zhipu Z.ai: GLM-4.6

    Zhipu Z.ai 7/13 GLM-4.6: AutoGLM-Coder refresh + agent scaffold, commercial-only on BigModel

    Zhipu Z.ai BigModel platform news on 2026-07-13 publishes a GLM-4.6 refresh with AutoGLM-Coder coding-specialist post-training + agent scaffold. Self-reported figures: MMLU-Pro 81.2, GPQA-Diamond 73.5, HumanEval+ 90.8, MultiPL-E 87.6, SWE-Bench Verified 67.8, Terminal-Bench 2.1 74.6. License: commercial-only via BigModel / Z.ai; open-source variants will follow separately (vendor-stated future commitment, not a current artifact). Semantic boundary: vendor platform news + Hugging Face mirror card; benchmark numbers are vendor-supplied.

    1 daily topicOriginally published Jul 13, 2026View source
  • Jul 14, 2026

    media

    aihot.virxact.com: meta-louisiana-5gw-datacenter-50b-investment

    Meta 7/13 Louisiana data center to 5GW / >$50B total + Entergy gas-fired + battery + nuclear expansion (per IT之家)

    aihot 2026-07-13 18:35 + IT之家: Meta expanding Louisiana data center to 5GW with >$50B total investment; commits to bearing all local energy + water costs + >$1B local infrastructure; signed Entergy deal for new gas-fired + battery-storage + nuclear capacity additions.

    1 daily topicOriginally published Jul 13, 2026View source
  • Jul 14, 2026

    media

    aihot.virxact.com: xai-grok-cli-silent-codebase-upload-disclosure

    xAI Grok CLI 7/13 silent-upload disclosure (per buzzing.cc HN translation)

    aihot 2026-07-13 08:10 + buzzing.cc: xAI's official Grok CLI v0.2.93 silently uploads working directory + ~/.claude.json + global AGENTS rules + 30+ Skill files + an API key on every turn via a side channel to a xAI Google Cloud bucket; xAI subsequently added a server-side toggle to disable codebase upload.

    1 daily topicOriginally published Jul 13, 2026View source
  • Jul 14, 2026

    media

    aihot.virxact.com: ploy-ai-agent-migrated-claude-opus-to-gpt-5-6-sol

    Ploy 7/13 AI-agent migration: Claude Opus 4.8 → GPT-5.6 Sol (per buzzing.cc HN)

    aihot 2026-07-13 07:54 + buzzing.cc: Ploy migrated AI-agent default from Claude Opus 4.8 to GPT-5.6 Sol — build time 2.2x faster, cost -27%, output tokens halved, visual score +0.034 — but GPT-5.6 Sol silently fills defaults on 25 tool params, leaving 52-64% of file reads empty.

    1 daily topicOriginally published Jul 13, 2026View source
  • Jul 14, 2026

    media

    aihot.virxact.com: altman-amodei-ai-net-jobs-narrative-shift

    Altman + Amodei 7/12 soften 'AI replaces jobs' framing (per The Decoder)

    aihot 2026-07-12 17:28 + The Decoder: OpenAI CEO Sam Altman + Anthropic CEO Dario Amodei both softened earlier 'AI replaces jobs' framings; Altman 'pretty confident' AI has so far net-created jobs; cross-referenced studies find no aggregate productivity / labor-market impacts.

    1 daily topicOriginally published Jul 12, 2026View source
  • Jul 14, 2026

    media

    aihot.virxact.com: apple-v-openai-trade-secrets-ai-hardware-suit

    Apple v. OpenAI 7/11+ trade-secrets suit (per IT之家 + TechCrunch AI)

    aihot 2026-07-11/12 + IT之家 + TechCrunch AI: Apple filed trade-secrets suit in U.S. NDCA against OpenAI (and Tang Tan + Chang Liu + io Products); ~400 ex-Apple employees now at OpenAI; consumer AI hardware at stake; February 2026 private-settlement attempt declined by OpenAI.

    1 daily topicOriginally published Jul 11, 2026View source