2026-09-26 AI 日报
统计日时区:America/Los_Angeles · 尚未结算
mobile-next/mobile-mcp · pacifio/atlas · dream-num/univer
开源项目
- mobile-next/mobile-mcp
Model Context Protocol Server for Mobile Automation and Scraping (iOS, Android, Emulators, Simulators and Real Devices)
- pacifio/atlas
Source control for agents. Use multiple coding agents, track their changes and query them in one place
- dream-num/univer
The Office Harness for AI Agents — Spreadsheets, Docs, Slides, Canvas, Relational Tables, and PDF in one runtime.
- rohitg00/ai-engineering-from-scratch
Learn it. Build it. Ship it for others.
- mvschwarz/openrig
Multi-agent harness that runs Claude Code and Codex together as one system
- paperclipai/paperclip
The open-source app everyone uses to manage agents at work
- heygen-com/hyperframes
Write HTML. Render video. Built for agents.
- melgarafael/DeskcommCRM
Open-source AI sales OS — self-hosted CRM with native AI agents + WhatsApp (WAHA). Open alternative to Kommo, Octadesk & Intercom for any business that sells…
- NVIDIA/Model-Optimizer
A unified library of SOTA model optimization techniques like quantization, distillation, pruning, neural architecture search, speculative decoding, etc. It…
- MakazhanAlpamys/Soup
Fine-tune LLMs from one YAML. Layer streaming trains an 8B model on a 4 GB laptop GPU.
论文
本统计日暂无条目。
行业新闻
- Unsealed Briefs in Authors’ Case v. Microsoft/OpenAI
- DeepSeek Elastic Compute (DSec)
arXivLabs is a framework that allows collaborators to develop and share new arXiv features directly on our website. Both individuals and organizations that…
- How to keep enjoying programming in a world of LLMs
- OpenAI bots meddled with multiple US Government agency sites
OpenAI has acknowledged that it alerted "dozens" of global institutions that their websites may have been meddled with by its AI bots acting improperly. AI…
- CEO of Mistral: AI is software. It can be controlled
- OpenAI agents tried to bruteforce a UN website's API fields
- Turning GLM-5.3-Flash into a Jev-like decision model
We found an approach to get Jev-like properties from standard LLMs like GLM-5.3-Flash. The core idea is to craft the input prompt so that the first output…
- OpenAI Codex agents go rogue and consumes USD 78,000 without authorization
My OpenAI CODEX account went rogue and from a simple request took the autonomous decision to launch 826 parallel agents / threads without any authorization on…
- The Copilot+ PC brand is dead
- Understanding the Impact of LLM Watermarking on AI Agent Behavior
社区动态
- 42x Faster Prompt Lookup Drafting in llama.cpp
- OpenAI stopped all frontier training, evaluation, and inference with tool-use (defined broadly) on the 20th of September and they are not resuming any of these activities for now
Discovery: Sep 20, 2026 Report updated: Sep 25, 2026 "An agent attempting to complete a search-based training task queried a public chatbot service through a…
- A Little Guide to Learning Distributed Algorithms for LLMS Training and Inference [D]
Distributed Training and Inference both involves having a fundamental understanding of how distributed systems work in general Distributed Parallelism Tensor…
- Swift 1.5 27b: Swift Qwen just got faster
Enjoy! Fucking loving it.
- LLMs were told they could lie in Diplomacy. Here's who actually kept their promises. [D]
the stats are from the game of diplomacy. diplomacy is basically a strategy game where you negotiate, form alliances, betray, and outmaneuver other players to…
- Qwen 3.8 flash next is based on Qwen 4 architecture, if the announced Qwen 4 27b is also the same architecture with n-grams does it mean I can actually have faster inference on a single 3090 without tweaking much?
I wish Qwen also released dataset and method to fully train a model ourselves but it is what it is. However, I come here with my stupid question because…
- Splash 1.1.0 released, GGUF quants support, MLX import and more
On my M5 Pro 64GB I can comfortably work in an agentic setup with the Qwen3.8 27B model in good quality (Unsloth UD-Q4_K_XL) at a decent speed of 50 t/s.…
- Introducing KoboldCpp Agent (and a plea for help)
Hello r/localllama once again, it's me your kobold concedo Been a few months since I last posted here, and today I have something new I'd like to share.…
- Teaching Neural Nets to Fight with RL [P]
In this project I wanted to see if any interesting emergent behaviors would appear if we trained two agents to play a streetfighter-like game using RL. Maybe…
- We have implanted 100 facts into the engram table of Qwen 3.8 Flash Next, and we have now created a website to explain it.
We’ve created an “engraft-engram” method to implant new facts into the Qwen 3.8 Flash Next engram table (and in the future, DeepSeek v4.1 Flash as well)…
官方发布
- Introducing Grok 4.7
- Qwen-Image-2.1: Compact, Efficient, and Unified Image Creation
We are excited to open-source Qwen-Image-2.1, an image model in the Qwen family that balances generation quality, inference efficiency, and cost.…
- Introducing GPT-6 Sol and Luna
Meet GPT-6 Sol and Luna, two models that bring frontier intelligence to everyday work with different balances of capability and cost.
- Coding sessions are longer and use more context. Claude Opus 5.5 is built with that in mind.
- Qwen/Qwen-Image-2.1-PE-T2I
text-to-image
- Yes, Claude can do Nine Loops
Guest writer and physicist Matt von Hippel shares what happened when he issued a challenge to AI companies to solve a problem in his former subfield of…
- Better prompt caching for GPT-6
Learn how GPT-6 improves prompt caching with higher cache hit rates, new diagnostics, explicit breakpoints, and controls that reduce latency and costs.
- Gemini 3.8 text-to-speech says hello
Gemini 3.8 Flash-Lite TTS and Gemini 3.8 Flash TTS are our most expressive audio models yet.
- Claude discovers a novel enzyme system with CRISPR-like repeats
We’re announcing a new life sciences research group and laboratory at Anthropic. This post introduces the team behind this work and shares early results in…
- Introducing MentalHealthBench
MentalHealthBench is an expert-informed benchmark for evaluating helpful and safe AI responses across realistic mental health conversations.