AI Resonance · 阅读最新日报 · AI 入门推荐

2026-10-02 AI 日报

统计日时区:America/Los_Angeles · 尚未结算

Niko1221/Strata · zeronsh/zeron · Panniantong/Agent-Reach

在应用中阅读本期

开源项目

  1. Niko1221/Strata

    Qwen3.8-Flash-Next on any consumer hardware: one-click install for Windows / Linux. Strata inference engine, OpenAI/Anthropic API on localhost, optional image…

  2. zeronsh/zeron

    A native control plane for Claude Code, Codex, Cursor, Devin and other coding agents.

  3. Panniantong/Agent-Reach

    Give your AI agent eyes to see the entire internet. Read & search Twitter, Reddit, YouTube, GitHub, Bilibili, XiaoHongShu — one CLI, zero API fees.

  4. DietrichGebert/ponytail

    Makes your AI agent think like the laziest senior dev in the room. The best code is the code you never wrote.

  5. mvschwarz/openrig

    Build your own network of agents from Claude Code, Codex and Pi: persistent teams with roles, shared context and owned work.

  6. colbymchenry/codegraph

    Pre-indexed code knowledge graph, auto syncs on code changes, for Claude Code, Codex, Gemini, Cursor, OpenCode, AntiGravity, Kiro, CoPilot, and Hermes Agent —…

  7. calesthio/OpenMontage

    World's first open-source, agentic video production system. 12 production pipelines, 100+ tools, 700+ agent skill and production-knowledge files. Turn your AI…

  8. p-e-w/heretic

    Fully automatic censorship removal for language models

  9. NVIDIA/SkillSpector

    Security scanner for AI agent skills. Detect vulnerabilities, malicious patterns, security risks, prompt injection, data exfiltration, and supply-chain risks…

  10. JuliusBrussee/caveman

    🪨 why use many token when few token do trick. Viral skill + proxy for coding agents that cuts 65% of tokens by talking like a caveman.

论文

  1. OneStreamer: Unifying Perception, Memory, and Proactive Response in Streaming Video Interaction

    Streaming video LLMs must retain evidence before its relevance to future tasks is known and respond when sufficient evidence becomes available. The challenge…

  2. Hierarchical Continuous Diffusion Language Models

    Discrete diffusion language models offer a compelling alternative to autoregressive generation for tasks demanding bidirectional reasoning and global…

  3. Beyond Memory: Harnessing Long-Horizon Agents with Explicit Belief States

    Large language model (LLM) agents can now undertake increasingly complex tasks, but the way they organize interaction history into memory does not ensure a…

  4. World Observer: Joint Actor-Observer Generation for Persistent World Modeling

    How can a world model continuously observe regions beyond the actor's current view? Video world models simulate how an environment evolves from an agent's…

  5. Sharpening Tax in Post-Training

    An emerging hypothesis about reinforcement learning (RL) post-training of large language models (LLMs) is that it merely sharpens existing behaviors of a base…

  6. Video Generation Models: A Survey of Post-Training and Alignment

    Video generation has rapidly progressed from short, low-quality clips to high-resolution, long-duration sequences with complex spatiotemporal dynamics.…

  7. ROWBench: Do Video Models Render What the Program Specifies?

    Programmable world models separate executable dynamics from visual generation, offering a promising foundation for next-generation game engines. However,…

  8. On-Policy or Off-Policy Learning? A Systematic Study of Distillation Dynamics

    On-policy learning has been argued to reduce catastrophic forgetting, produce sparser parameter updates, and improve generalisation. However, existing…

  9. Adaptive Reward Routing: Dynamic Multi-Reward Optimization for Joint Audio-Video Diffusion via Forward-Process RL

    Multi-reward guided reinforcement learning (i.e., RL) offers a promising way to improve joint audio-video diffusion models along several complementary…

  10. Agent Priors-guided Policy Learning

    Robots that learn from a few demonstrations often require two forms of generalization. Compositional generalization recombines skills to solve new tasks, and…

行业新闻

  1. From the creator of Redis; run LLM locally with ds4

  2. One month coding with GLM 5.3 Flash

  3. The Four Horsemen of Agentic Coding

  4. Don't be fooled–LLMs don't reason

  5. GPT-6 Astra plays World of Warcraft for the first time with agent-wow

  6. Claude-Shaped Science

    Summary: In this guest post, Prof. Matthew Schwartz returns to describe a new approach to AI-accelerated science. In Vibe Physics, Schwartz discussed…

  7. Harvard particle physicist Matthew Schwartz drops 36 papers authored with Claude

  8. Decision models like Jev don't beat LLM-as-a-judge or traditional classifiers

  9. Ask HN: Is anybody producing good code with coding agents?

    This is a genuine problem that I hear from senior engineers. I'm looking for a solution. -- The quality of ai-generated code is <censored> (claude, agy,…

  10. Apple is tightening macOS 'Full Disk Access' due to new risks from AI agents

社区动态

  1. Project Suncatcher has liftoff!

    Project Suncatcher has liftoff! ☀️🚀 Announced last year, this moonshot project explores if we can one day host machine learning infrastructure in space.…

  2. qwen4exp: halve the indexer score memory by ServeurpersoCom · Pull Request #29825 · ggml-org/llama.cpp

    Qwen Flash Next now uses less VRAM

  3. Try Grok 4.7 in Ramp Router, now 50% off until October 6th

    Try Grok 4.7 in Ramp Router, now 50% off until October 6th

  4. llama, server: add /v1/systemone API (models: laya, julia-1, lev, openjev, kev) by ngxson · Pull Request #29818 · ggml-org/llama.cpp

    article: https://huggingface.co/blog/ggml-org/decision-models-in-llamacpp Now you can jev without jev https://huggingface.co/ggml-org/OpenJev-GGUF…

  5. I made my iPhone a second GPU for my 24 GB MacBook: Qwen 3.8 27B prefills 29–44% faster & my holds part of the CTX window.

    **DISCLAIMER** THE PREFILLING TPS SHOWN ON THE PHONE IS COMPUTED ONLY FOR THE LAYERS IT HOLDS. ALREADY FIXING IT TO SHOW END-TO-END PREFILL RATE. NUMBERS…

  6. LibLayaX: run the Laya AI decision model inside your own app

    Dear LLM developers community, I have just released four open-source projects today that let an application use the Laya model directly, with no server and no…

  7. Gemini 4 Argon has been added to the Gemini API docs.

    Gemini 4 Argon has been added to the Gemini API docs. Hopefully we might see a public release upcoming week.

  8. Topological Out-of-Domain Generalization in Dynamical Systems Reconstruction [R]

    In our #NeurIPS2026 paper “ Topological Out-of-Domain Generalization in Dynamical Systems Reconstruction ” (preprint: https://arxiv.org/abs/2606.22969) we try…

  9. [R] Would you keep a robot demonstration if hand tracking missed the moment the plug went in?

    Suppose you’re recording a human plugging a cable into a socket to collect demonstrations for robot learning. The hand tracker captures the approach…

  10. Evaluating 14 LLMs as a visual coding agent: re-rendering, a deterministic judge, and the provider quirks that skewed my first results

    I built a bench for a photo-to-Blender agent (the model writes and runs Blender Python to rebuild a photo as an editable scene) and ran 14 models through it.…

官方发布

  1. Gemini 4 Argon: our next era of frontier intelligence

    Announcing Gemini 4 Argon, our frontier model for real-world coding, enterprise knowledge work, and cyber defense, rolling out soon.

  2. Introducing GPT-6.1 Sol

    Meet GPT-6.1 Sol: near-Astra intelligence for coding, computer use, and professional work at one-fifth of Astra’s standard API input and output token prices.

  3. grok-voice-transcribe-1.0 end of life

    grok-voice-transcribe-1.0 is deprecated and reaches end of life on October 2, 2026. All requests to that slug are routed to grok-voice-transcribe-2.0 at the…

  4. We've launched Claude Sonnet 5.5 (claude-sonnet-5-5).

    We've launched Claude Sonnet 5.5 (claude-sonnet-5-5). It's available on the Claude API, Claude in Amazon Bedrock, Claude Platform on AWS, Claude on Google…

  5. Basis completes a tax workbook 2x faster with GPT-6 Astra

    GPT-6 Astra completed a 50-tab tax workbook twice as fast as GPT-5.6 Sol, and its stronger understanding of user intent gives Basis more confidence in…

  6. Introducing dots

    Dots by OpenAI are proactive assistants that can keep working across complex projects and everyday tasks. Learn how dots help you stay in control while work…

  7. Anthropic invests $100 million to train 10,000 engineers and tackle the enterprise AI talent gap

    Anthropic is investing $100 million in Claude Frontier Academy to train 10,000 Frontier Deployed Engineers by the end of 2027, with cohorts from Accenture,…

  8. Z.ai GLM 5.3 (zai-glm-5-3) is now Generally Available.

    Z.ai GLM 5.3 (zai-glm-5-3) is now Generally Available.

  9. Claude-shaped science

    Guest author Prof. Matthew Schwartz describes what happened when he stopped fighting Claude and allowed Claude to find “Claude-shaped” problems: ones best…

  10. The latest AI news we announced in September 2026

    Here are Google’s latest AI updates from September 2026