AI Resonance · 阅读最新日报 · AI 入门推荐

2026-09-29 AI 日报

统计日时区:America/Los_Angeles · 已结算

NVIDIA/OpenShell · VectifyAI/PageIndex · mvschwarz/openrig

在应用中阅读本期

开源项目

  1. NVIDIA/OpenShell

    OpenShell is the safe, private runtime for autonomous AI agents.

  2. VectifyAI/PageIndex

    📑 PageIndex: Document Index for Vectorless, Reasoning-based RAG

  3. mvschwarz/openrig

    Multi-agent harness that runs Claude Code and Codex together as one system

  4. t8y2/dbx

    25 MB lightweight cross-platform database client for 100+ databases, including MySQL, PostgreSQL, SQLite, Redis, MongoDB, DuckDB, SQL Server, and Dameng.…

  5. DietrichGebert/ponytail

    Makes your AI agent think like the laziest senior dev in the room. The best code is the code you never wrote.

  6. firecrawl/firecrawl

    🔥 Supercharge your AI agents with data from the web and beyond. A web data API to search, scrape, and access more sources.

  7. rohitg00/ai-engineering-from-scratch

    Learn it. Build it. Ship it for others.

  8. ifixai-ai/iFixAi

    Independent Auditing of AI Agents. Run by human or the agent itself, to answer the most crucial question in the AI Agent Economy. Is the agent doing what is…

  9. TencentCloud/Octop

    A smarter, self-hosted AI assistant — multi-user, multi-agent.

  10. mattpocock/skills

    Skills for Real Engineers. Straight from my .agents directory.

论文

  1. YuE2: Unifying Symbolic and Audio Music Generation at Frontier Quality

    Symbolic models make melody, harmony, rhythm, and form explicit but typically stop before a finished recording; audio models produce complete songs while…

  2. MassAlloc Attention: Let Attention Allocate Its Own Compute

    FullAttn often assigns negligible normalized mass to much of the causal score space, yet dense kernels execute the complete post-score path after forming each…

  3. Post-Training Leaves Behavioral Shadows on Unrelated Decisions

    We find that language models can transfer capabilities through task-unrelated text. Post-training typically improves language models using task-specific data.…

  4. Self-Evolving Coding Agents: From Digital Programs to Physical-World Intelligence

    Vision-language-action (VLA) and world-action (WAM) models map observations and instructions directly to robot actions. This directness ties a policy to…

  5. Beyond Teacher Assignment: Domain-Normalized Multi-Teacher On-Policy Distillation

    Reinforcement learning can turn one language model into several specialists, each excellent at a single skill such as mathematics, coding or following…

  6. Improving Test-Time Scaling with Adaptive Looped Transformers

    Looped transformers have demonstrated promising parameter efficiency by reusing layers for latent computation. Prior studies compare looped and non-looped…

  7. TraceDance: An Automated System for Building Agent Behavior Benchmarks from Real-World Agent Deployment Traces

    An agent can complete a task while exhibiting undesirable behavior during execution. Developers need tests for the specific behaviors encountered in…

  8. Groupwise Agentic Grading and Advantage Redistribution for Code Agent RL

    Reinforcement learning (RL) for code agents often uses executable tests to provide binary rewards. With these rewards, Group Relative Policy Optimization…

  9. Learning Native Reflection in Unified Models with Interleaved Reinforcement Learning

    Unified multimodal models can both look at and render images, so in principle they can repair their own generations: diagnose what an image gets wrong, revise…

  10. Duplex-MPE: Benchmarking Multi-Party Interaction in Full-Duplex Dialogue

    Real-time full-duplex speech models can listen while speaking, enabling natural interaction without rigid turn boundaries. Existing benchmarks evaluate…

行业新闻

  1. GPT 6.1 Sol: Near-Astra intelligence for a fifth of the price

  2. Dots: Always-on agents

  3. GLM-5.3 and the spread of advanced cyber capabilities

    Andrew Fasano, Marius Fleischer Cole McFaul, Robert Xiao, Tripp Gallagher Five months ago, we announced Claude Mythos Preview, the first AI model that could…

  4. A Privacy Analysis of Web and Mobile Conversational AI Agents [pdf]

  5. Jeeves. Reasoning improves Jev-like decision models

  6. ChatGPT Pro 500

  7. Unsurprisingly, Meta's new Muse AI agent blatantly ignores users permissions

  8. Claude partial outage

  9. DevDay 2026 Recap

  10. Language models for text classification: From bag-of-words to Jev

社区动态

  1. Ultrafast is our fastest way to build with Astra yet: in Codex, it runs up to 8x faster than Astra Standard and 4x…

    Ultrafast is our fastest way to build with Astra yet: in Codex, it runs up to 8x faster than Astra Standard and 4x faster than Astra Fast. Bring your ideas to…

  2. Anthropic just dropped the greatest advertisement for GLM ever.

    Like.. yea bro, I knew GLM was cool. Now everyone does.

  3. OpenAI launches GPT-6.1 Sol

    "Near-Astra intelligence for a fifth of the price"

  4. What do you want from AI?

    What do you want from AI? We’re launching a new study with Anthropic Interviewer to learn more about your experiences using AI, what role you want it to play…

  5. GLM-5.3 and the Spread of Advanced Cyber Capabilities \ Anthropic

  6. Qwen3.8 flash next ISTA-DASLab GGUF 50t/s TG and 1500t/s PP with 12GB VRAM and 64GB RAM Laptop on 'Strata' engine

    I think most people are sleeping on this inference engine. I tried multiple llama.cpp forks and none of them comes close to the inference speed of Strata.…

  7. I wrote a free, open-source book on making ML models actually fast, from silicon to agents [P]

    I’ve spent the last few months writing something I wish I had when I started working on ML performance engineering. It’s called How to Make Your Model Fast: A…

  8. Changed one environment variable and our ticket tagging was down for most of the morning

    We have a small feature that tags incoming support tickets before a person sees them, mostly deciding whether something is a billing problem or a bug report.…

  9. I built a little eBPF tool to see what my coding agents actually do on my machine

    I've been running Claude Code, Codex, ohmypi, and a couple others locally and it kind of bugged me that I had no real idea what they were doing outside the…

  10. No cameras were harmed in the making of this Odyssey pilot.

    No cameras were harmed in the making of this Odyssey pilot.

官方发布

  1. Introducing GPT-6.1 Sol

    Meet GPT-6.1 Sol: near-Astra intelligence for coding, computer use, and professional work at one-fifth of Astra’s standard API input and output token prices.

  2. Introducing dots

    Dots by OpenAI are proactive assistants that can keep working across complex projects and everyday tasks. Learn how dots help you stay in control while work…

  3. GLM-5.3 and the spread of advanced cyber capabilities

    Like Claude Mythos Preview, GLM-5.3 has strong capabilities for autonomously building end-to-end cyber exploits. But GLM-5.3 is unlike other frontier models…

  4. Z.ai GLM 5.3 (zai-glm-5-3) is now Generally Available.

    Z.ai GLM 5.3 (zai-glm-5-3) is now Generally Available.

  5. deepseek-ai/dsh-libreoffice-kit

    An internal component used by DeepSeek Harness

  6. deepseek-ai/dsh-node-addon-require-builtin

    An internal component used by DeepSeek Harness

  7. Coding sessions are longer and use more context. Claude Opus 5.5 is built with that in mind.

  8. What do you want from AI?

    We’re launching a new study using Anthropic Interviewer to learn from your experiences with AI, and we invite you to participate.

  9. DevDay 2026 Recap

    Explore more than 20 announcements from OpenAI DevDay 2026, including GPT-6 Astra, ChatGPT, Codex, APIs, security, and new tools for builders.

  10. deepseek-ai/DeepEP-Ascend

    A high-performance communication library for machine learning training and inference on Huawei Ascend NPUs.