AI Resonance · 阅读最新日报 · AI 入门推荐

2026-09-23 AI 日报

统计日时区:America/Los_Angeles · 已结算

Nasiko-Labs/nasiko · strands-agents/harness-sdk · farion1231/cc-switch

在应用中阅读本期

开源项目

  1. Nasiko-Labs/nasiko

    Developer Control Plane for your AI Agents

  2. strands-agents/harness-sdk

    Build an agent harness and control it end-to-end. Open-source SDK for production AI agents in Python & TypeScript - any model, any cloud.

  3. farion1231/cc-switch

    A cross-platform desktop All-in-One assistant for Claude Code, Codex, OpenCode, OpenClaw, Grok Build & Hermes Agent. Only official website: ccswitch.io

  4. experientiallabs/experiential

    Experiential is the open source, zero markup gateway for BYOK, self-hosted and 1000+ marketplace models. It learns from your traffic to cut costs, recommend…

  5. google/ax

    Google's open agentic orchestration runtime

  6. aayushch/laya

    Laya is an open-source, local-first AI notification command center that aggregates Slack, Gmail, GitHub, Jira, Notion, Outlook, Calendar (and more)…

  7. dream-num/univer

    The Office Harness for AI Agents — Spreadsheets, Docs, Slides, Canvas, Relational Tables, and PDF in one runtime.

  8. every-app/open-seo

    Open source alternative to Semrush and Ahrefs

  9. TNT-Likely/PanWatch

    盯盘侠 PanWatch · 自托管 AI 盯盘助手,集成 TradingAgents 多 Agent 投资决策 | A股/港股/美股实时监控、持仓管理、智能分析、全渠道推送

  10. rohitg00/ai-engineering-from-scratch

    Learn it. Build it. Ship it for others.

论文

  1. GAE: Learning a Geometry-Native Latent Space for 3D-Consistent World Generation

    We present a compact geometry-native latent space as a shared foundation for perception and generation. Visual generators can produce photorealistic frames…

  2. The Tasteful Agent: Measuring and Improving Taste in Long-Horizon Tasks

    LLM agents increasingly work on long-horizon tasks, and the decisions they make along the way, such as which hypothesis to test or which implementation to…

  3. Ovis-Embedding: Pushing the Frontiers of Universal Omni-Modal Embeddings

    In this report, we introduce Ovis-Embedding, a state-of-the-art omni-modal embedding family built on native integration of text, image, video, and audio.…

  4. All-in-One Multilingual Scene Text Recognition with Script-aware Mixture-of-Experts

    Multilingual scene text recognition (STR) remains challenging due to the scarcity of training data for most languages and the difficulty of serving diverse…

  5. RULER: Instance-aware Rubric Rewards for SVG Generation

    Generating Scalable Vector Graphics (SVG) code from natural-language instructions is an open-ended task without absolute visual ground truth, leaving both…

  6. From Pattern Recognizers to Personalized Companions: A Survey of Large Language Models in Mental Health

    The rising global prevalence of mental health conditions, together with longstanding barriers in traditional healthcare, such as limited resources, high cost,…

  7. Lean Pool: An AI-Maintained Archive of Formalized Mathematics

    Lean Pool is a repository of formalized mathematics. It is grown, maintained and optimized by AI agents.

  8. Bellman Policy Optimization

    Reinforcement learning with verifiable rewards (RLVR) improves the reasoning capabilities of large language models (LLMs). We introduce Bellman Policy…

  9. Flash-dLLM: IO-Aware KV Caching and Parallel Decoding for Fast, Memory-Efficient Diffusion LLMs

    Diffusion Large Language Models (dLLMs) have recently emerged as a promising alternative to autoregressive LLMs by enabling non-autoregressive text…

  10. StableVQ: Practical Guidelines for Stable Vector-Quantized Tokenizer Training

    Vector Quantization (VQ) is fundamental to discrete visual tokenizers that power modern autoregressive and masked image generation models. While recent…

行业新闻

  1. Claude discovers a novel enzyme system with CRISPR-like repeats

    We’re introducing a new life sciences research group and laboratory at Anthropic. Our focus is on fundamental biology research using Claude: exploring…

  2. Gemini 3.8 text-to-speech

    Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS are our most expressive audio generation models yet. Generate custom character voices and direct scene…

  3. Claude Code reads AGENTS.md only when telemetry is on [fixed]

  4. Feds Target AI Critics as "Foreign Agents"

  5. Early rogue AI agent activity and attempts to hack found on urlquery.net

  6. OpenAI agent hacked Australian government website, PM says

    We can bring you more now from Nick Clegg, who has been speaking to BBC Radio 4's Today programme. Asked whether concerns about uncontrolled AI development…

  7. GPT-6 Astra has gained the ability to drive a car

  8. OpenAI breaches Medicare, Albanese reveals

  9. Once Claude can measure something, it can make it faster

  10. OpenAI is enlisting an influencer army to make it look 'good for the world'

社区动态

  1. Claude has discovered a previously unknown enzyme system hidden in the DNA of bacteriophages.

    Claude has discovered a previously unknown enzyme system hidden in the DNA of bacteriophages. Beside the enzyme’s gene sits a long array of repeating DNA—a…

  2. introducing Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS, our most expressive audio generation models yet these…

    introducing Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS, our most expressive audio generation models yet these models enable creators, developers, and…

  3. ⚡ Meet Qwen-Audio-3.1!

    ⚡ Meet Qwen-Audio-3.1! ASR, TTS & Realtime are fully upgraded, joined by two new models: TTS-Next for audio creation and ASR-Next for audio understanding.…

  4. Introducing Qwen Intelligence, bringing personal intelligence within everyone's reach.

    Introducing Qwen Intelligence, bringing personal intelligence within everyone's reach. 📱✨ It launches with three SOTA agents: 🥳 - Mobile Planner Agent:…

  5. We heard you loud and clear.

    We heard you loud and clear. ChatGPT Voice can now: - Use plugins like your email, calendar, and Slack. - Be powered by GPT-6 Astra, Sol, and Luna. - Be used…

  6. Today we announced the Claude-led discovery of a molecular machine that we suspect could represent a new gene editing…

    Today we announced the Claude-led discovery of a molecular machine that we suspect could represent a new gene editing mechanism. Its precise function,…

  7. We’re demonstrating how frontier models have continued to improve in realistic mental health conversations with…

    We’re demonstrating how frontier models have continued to improve in realistic mental health conversations with MentalHealthBench. This new open benchmark was…

  8. Claude discovered a novel enzyme system with properties reminiscent of CRISPR

    Claude has discovered a previously unknown enzyme system hidden in the DNA of bacteriophages. Beside the enzyme’s gene sits a long array of repeating DNA—a…

  9. Introducing Support for Local AI Models in the Antigravity SDK

  10. OpenAI rolls out upgraded prompt caching and reduced cached input rates for GPT-6

官方发布

  1. Introducing GPT-6 Sol and Luna

    Meet GPT-6 Sol and Luna, two models that bring frontier intelligence to everyday work with different balances of capability and cost.

  2. Introducing Grok 4.7

  3. Qwen-Image-2.1: Compact, Efficient, and Unified Image Creation

    We are excited to open-source Qwen-Image-2.1, an image model in the Qwen family that balances generation quality, inference efficiency, and cost.…

  4. Gemini 3.8 text-to-speech says hello

    Gemini 3.8 Flash-Lite TTS and Gemini 3.8 Flash TTS are our most expressive audio models yet.

  5. Claude discovers a novel enzyme system with CRISPR-like repeats

    We’re announcing a new life sciences research group and laboratory at Anthropic. This post introduces the team behind this work and shares early results in…

  6. Better prompt caching for GPT-6

    Learn how GPT-6 improves prompt caching with higher cache hit rates, new diagnostics, explicit breakpoints, and controls that reduce latency and costs.

  7. On Claude Opus 5.5, thinking can't be disabled: thinking: {"type": "disabled"} and thinking: {"type": "enabled", ...}…

    On Claude Opus 5.5, thinking can't be disabled: thinking: {"type": "disabled"} and thinking: {"type": "enabled", ...} return a 400 error. Omit the thinking…

  8. Introducing Grok Voice Transcribe 2.0

  9. Introducing Support for Local AI Models in the Antigravity SDK

    The Google Antigravity SDK now empowers developers to execute offline, agentic workflows locally using models like Gemma 4 26B A4B via LiteRT. This update…

  10. Qwen/Qwen-Image-2.1-PE-T2I

    text-to-image