AI 每日快讯

AI 每日快讯

AI 产品、模型、开源工具和官方动态的时间流。保留历史记录,按分类、日期和标签继续筛选。

3779历史快讯
207开源工具
80当前结果
10 月 02 日 今日快讯
MarkTechPost 官方资讯

MarkTechPost:A Coding Guide to Google Research’s Kauldron: Configs That Are Plain Data, Components Wired …

原文摘要:In this comprehensive coding guide, we explore Google Research's Kauldron—a JAX training library optimized for research velocity and modularity. Learn how konfig turns experiments 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:Cloudflare Releases Clef and Clef-flash: Open-Weight Decision Models That Return Typed Proba…

原文摘要:Cloudflare has released Clef (27B) and Clef-flash (9B), open-weight decision models that return typed probabilities instead of text. They are Jev-API compatible, accept images, and 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

10 月 01 日 昨日快讯

Cohere 发布 Embed 5:分 Pro 与 Fast 两档,支持文本图像融合输入

Cohere 发布了新的嵌入模型家族 Embed 5,值得关注是因为嵌入模型直接决定企业搜索与 RAG 的检索质量。原始信息显示,Embed 5 面向企业搜索、RAG 和 Agent 检索,分为两档:Embed 5 Pro 追求最高检索质量,Embed 5 Fast 面向在线查询路径的延迟与成本;两者都接受文本、图像以及文本加图像的融合输入。它影响的是做企业知识库、多模态检索和 Agent 记忆层的团队。下一步验证方式是结合自身语料做召回对比测试,重点看融合输入场景下的检索准确率与查询延迟。

MarkTechPost 官方资讯

MarkTechPost:NVIDIA Releases Kumo Tabular: Open Tabular Foundation Models That Predict New Rows in a Sing…

原文摘要:NVIDIA has released Kumo Tabular, a new family of tabular foundation models (TFMs) for classification and regression. If you have followed TabPFN or TabICL, the setup will look fam 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:Perplexity Releases pplx-embed-v2-context-9b-preview: A Contextual Embedding Model That Retr…

原文摘要:Perplexity Research and turbopuffer have released pplx-embed-v2-context-9b-preview, a contextual embedding model for RAG pipelines. Each chunk is embedded with the full document in 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 30 日 2026-09-30 快讯
MarkTechPost 官方资讯

MarkTechPost:OpenAI Releases GPT-6.1 Sol: Near-Astra Coding and Computer Use at One-Fifth of Astra’s Toke…

原文摘要:OpenAI released GPT-6.1 Sol on September 29, 2026, an upgrade to GPT-6 Sol. It reaches near-Astra results on agentic coding, computer use and professional work at one-fifth of Astr 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:Google DeepMind Unveils Gemini 4 Argon with 1M Output Tokens for Coding, Knowledge Work and …

原文摘要:Gemini 4 Argon tops GPT-6 Astra and Claude Opus 5.5 on most 评测, but access remains gated today. The post Google DeepMind Unveils Gemini 4 Argon with 1M Output Tokens for Co 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:Perplexity Introduces Photon: A Rust-Based Retrieval Engine That Cuts p99 Latency From 800 m…

原文摘要:Perplexity has released Photon, an in-house retrieval and ranking engine written in Rust. It replaces an open-source engine Perplexity had forked for its AI-native search stack. Ph 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:NVIDIA Researchers Introduce Physis-Lang: Self-Evolving Physical Language That Lifts Cosmos …

原文摘要:Video world models can render convincing clips that still break physics. Butter spreads like paint. Balls pass through walls. A team from NVIDIA, MIT and the University of Oxford a 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:One Bad Prompt Took Down a Company’s Salesforce: RSA’s Jim Taylor on Agent ID and Taming the…

原文摘要:RSA launched Agent ID at The AI Conference in San Francisco. It's an agentic identity security platform for finance, government, healthcare, and critical infrastructure. It has 3 m 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 29 日 2026-09-29 快讯
MarkTechPost 官方资讯

MarkTechPost:Liquid AI Releases d1: A Decision Model That Returns Calibrated Probabilities With Zero Outp…

原文摘要:Liquid AI has released d1, a decision model built for structured choices instead of text generation. You give it context and a set of typed questions. It returns calibrated probabi 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:OpenAI Launches dots: Always-On GPT-6 Astra Agents That Work From Their Own Cloud Computers

原文摘要:OpenAI just introduced dots at their DevDay today. Dots are persistent AI agents powered by GPT-6 Astra. Each dot gets its own cloud computer and browser. It works across 4,000+ ap 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:Nebius Opens 2026 Physical AI Awards: Five $150K Compute Prizes, Nine Judges, and an October…

原文摘要:Nebius and NVIDIA are running the 2026 Physical AI Awards for startups with products in the field. Five category winners each get $150,000 in compute credits, joint promotion, exec 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:Google Research Open-Sources RRSI: AI Agents That Improve Their Own Harness Without Overfitt…

原文摘要:Google Cloud AI Research has open-sourced RRSI, a framework that lets LLM agents rewrite their own prompts, tools and memory while model weights stay frozen. It adds a leakage crit 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:H Company Releases Holo4: Open-Weight Computer-Use Models That Click, Code and Call Tools Ac…

原文摘要:H Company has released Holo4, a family of generalist computer-use models for AI agents. One set of weights clicks and types on screens. It also writes code and calls MCP or API too 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

阿里 Qwen 发布 Qwen-Audio-3.1-Realtime:全双工语音模型会思考、会调工具、会判断何时开口

阿里 Qwen 团队发布了 Qwen-Audio-3.1-Realtime,这是一个全双工语音模型,被训练成能推理、调用工具并自行决定何时说话。原始信息给出的数据是:在 τ-Voice 适配任务上成功率从 78.4% 提升到 82.0%,对背景人声的误回应从 73% 降到 13%,目前已作为 API 在 QwenCloud 上线。值得关注的是,「何时开口」这个判断能力直接关系到语音助手的自然度,误回应大幅下降意味着在嘈杂环境里更可用,而工具调用则让它不止于聊天。受影响的主要是做实时语音交互、智能客服和语音 agent 的开发者。下一步建议在 QwenCloud 上申请 API,用真实噪声环境和多轮打断场景测试其响应时机与工具调用稳定性。

MarkTechPost 官方资讯

MarkTechPost:Anthropic Releases Claude Sonnet 5.5: 70.6% on Terminal-Bench 4.0 at the Same $2/$10 Price

原文摘要:Anthropic has released Claude Sonnet 5.5, the second model in its Claude 5.5 family. It scores 70.6% on Terminal-Bench 4.0 and lands within 2 points of Opus 5.5 on GDPval-AA. It al 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 28 日 2026-09-28 快讯
MarkTechPost 官方资讯

MarkTechPost:NVIDIA Launches Open Agent Safety Platform: OpenShell Sandboxes Agents on Vera CPUs While Se…

原文摘要:NVIDIA has launched the Open Agent Safety Platform, an open reference design that enforces AI agent safety outside the agent itself. OpenShell, an Apache 2.0 runtime, sandboxes age 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

Fireworks AI 发布 Ember-1:后训练 Kimi K3,Token 用量约降四成

Fireworks AI 发布了 Ember-1,这是基于 Kimi K3 后训练的模型,思路是让模型学会生成更短的推理轨迹,而不是简单调低推理强度。官方报告在一次生产环境 A/B 测试中,每任务输出 Token 从 49.3K 降到 29.9K,减少约 40%,而评分基本不变。Ember-1 目前以 API-only 研究预览形式提供,定价与 Kimi K3 相同。值得关注的原因是,推理成本正成为 Agent 规模化落地的核心约束,通过后训练压缩推理长度比粗暴降档更可能保住质量。它影响的是高频调用推理模型、对延迟和成本敏感的开发者与平台团队。下一步建议在相同任务集上对比 Ember-1 与 Kimi K3 的输出长度、准确率和端到端延迟,确认 40% 的节省在自有场景中是否成立。

Google Research 提出 AI 视频联合导演:四个 Agent 框架做分钟级连贯长视频

Google Research 发布了一套面向长视频生成的 AI 联合导演方案,包含四个 Agent 框架,目标是把短片段扩展成连贯的分钟级故事。它明确针对多镜头 AI 视频流程中最常见的两类失败:身份漂移和级联错误,也就是角色外观在不同镜头间不一致,以及前序错误在后续生成中不断放大。值得关注的是,当前多数视频生成工具擅长出高质量单镜头,却难以维持长叙事的一致性,这套框架试图用 Agent 分工来补上导演层能力。受影响的是做 AI 短片、广告分镜和长内容自动化的团队。下一步建议找到原始论文或项目页,确认四个框架各自职责、评测指标和是否开源,再用自己的多镜头脚本做小规模复现验证。

09 月 27 日 2026-09-27 快讯
MarkTechPost 官方资讯

MarkTechPost:A Coding Guide to Google Research’s MSEB: Writing Sound Encoders to the 评测 Contract a…

原文摘要:A comprehensive coding tutorial on Google Research's Massive Sound Embedding 评测 (MSEB), demonstrating how to implement custom sound encoders, drive classification, clusterin 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

企业级编码 Agent 对比:赔偿条款、数据驻留与 500 席位成本

MarkTechPost 梳理了 GitHub Copilot、AWS Kiro、Cursor、Devin 和 Windsurf 的合同条款,重点比较生成代码的知识产权赔偿、提示词存储、审计日志、数据驻留以及 500 席位的真实成本。文中提到 Copilot 和 Kiro 对生成代码提供不设上限的赔偿,而 Cognition 的标准条款完全排除输出相关责任。值得关注的是采购决策往往卡在法务而非功能,赔偿范围和数据存放位置可能直接决定能否过审。影响对象是准备大规模采购编码 Agent 的企业工程与法务团队。验证时建议直接调取各家最新合同原文,核对赔偿例外、提示词保留期限和区域选项,再按 500 席位口径重算总成本。

09 月 26 日 2026-09-26 快讯
MarkTechPost 官方资讯

MarkTechPost:Sarvam AI Releases Saaras V4: A Speech-to-Text Model for All 22 Indian Languages and Global …

原文摘要:Sarvam AI's Saaras V4 is a speech-to-text model covering all 22 Indian languages plus global English. It pairs an audio encoder with a 3B hybrid state-space decoder. It adds keyter 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:Supersonic Labs Releases Julia 1: A 144.3M-Parameter Open Decision Model That Runs on a CPU

原文摘要:Supersonic Labs has released Julia 1, a 144.3M-parameter decision model built on mmBERT-small. It takes context, a question, and 2 to 20 options, then returns one choice with proba 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:Exa Launches Agent Ultra: A Subagent Swarm Deep Research API Built for Exhaustive List Build…

原文摘要:Exa has released Agent Ultra, the highest effort mode of its Exa Agent API. It coordinates subagents across thousands of sources for list building and entity enrichment. Exa report 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 25 日 2026-09-25 快讯
MarkTechPost 官方资讯

MarkTechPost:Liquid AI Releases LFM2.5-VL-3B-DSpark: Speculative Decoding for Vision-Language Models With…

原文摘要:Liquid AI has released LFM2.5-VL-3B-DSpark, a 279.5M-parameter draft model that brings speculative decoding to its LFM2.5-VL-3B vision-language model. It delivers up to 3.13x faste 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:Aikido Security Releases Altar-1: An Open-Weight Security Model Pruned From GLM-5.3 to 328 G…

原文摘要:Aikido Security has released Altar-1, its first open-weight security model. It is a compressed version of Z.AI’s GLM-5.3, built to run inside infrastructure the customer controls. 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:Perplexity Trains Its Computer Agent on Real Mistakes With Hint-Guided Self-Distillation

原文摘要:Perplexity Research published a new post-training study. It trains a model inside Perplexity Computer on real user sessions, including failed ones. The method pairs rejection sampl 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:Fastino Releases GLiNER2.5-Decide: A 340M Open-Weight Decision Model That Runs on CPU

原文摘要:Fastino Labs has released GLiNER2.5-Decide, a 340M-parameter open-weight decision model. It takes text and a schema of typed questions and returns structured answers. Each answer c 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:Black Forest Labs Releases FLUX 3 Action: A 7B Open-Weights World Action Model That Tops Rob…

原文摘要:Black Forest Labs (BFL), the lab behind the FLUX image models, has released FLUX 3 Action. It is a 7B open-weights World Action Model (WAM) for robot control. The model reads camer 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 24 日 2026-09-24 快讯
MarkTechPost 官方资讯

MarkTechPost:BottleCap AI Releases ThinkingCap-Qwen3.8-27B: 37.2% Fewer Thinking Tokens at a 0.86pp Accur…

原文摘要:BottleCap AI has released ThinkingCap-Qwen3.8-27B, a fine-tune of Qwen3.8-27B that spends 37.2% fewer thinking tokens across 12 评测. Macro accuracy moves from 86.65% to 85.7 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:Contrastive-LM Releases CLM-8B: An Open System One Model That Scores Agent Actions Up to 9× …

原文摘要:Contrastive-LM has released CLM-8B, an open System One model that scores candidate actions against a state instead of generating text. It adds 2 small projection heads to a frozen 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:A Coding Guide to TypeSafe AI Jev: Typed Decisions, Calibrated Confidence, and Speculative F…

原文摘要:This tutorial provides a complete coding guide to TypeSafe AI's Jev, a System One model designed for non-text, structured judgments. It covers installing the official Python SDK, u 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 23 日 2026-09-23 快讯
MarkTechPost 官方资讯

MarkTechPost:Google Releases Gemini 3.8 Flash TTS and Flash-Lite TTS With Prompt-Based Voice Design

原文摘要:Google has released Gemini 3.8 Flash TTS and Flash-Lite TTS, 2 new text-to-speech models available now through the Gemini API and Google AI Studio. Flash TTS designs new voices fro 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:NVIDIA Releases Nemotron 3 Diarization: A 100M-Parameter Open-Weight Model That Tracks 8 Spe…

原文摘要:NVIDIA has released Nemotron 3 Diarization, an open-weight speaker diarization model on Hugging Face. It answers one question about any conversation: who spoke when. The 100M-param 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:Nokia Open-Sources AnyJev: A Training-Free Layer That Turns Any Open LLM Into a Calibrated D…

原文摘要:Nokia’s applied research team has open-sourced AnyJev, a Python library that turns an open LLM into a decision model. It needs no training. It targets a common production job: pick 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:Kyutai Releases Voice of Reason: A Speech-Native Model that Solves Spoken Math with Reinforc…

原文摘要:Kyutai has released Voice of Reason, 2 open-weight speech-to-speech models built on GLM-4-Voice-9B. Supervised fine-tuning and reinforcement learning lift spoken GSM8K accuracy fro 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:SpeakON Ships a MagSafe AI Voice Button With Its Own Microphone: Turning Your Voice into Pol…

原文摘要:Voice input on phones has been solved for years. What has not been solved is the output. Speak into most dictation tools and you get back exactly what you said, fillers and false s 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 22 日 2026-09-22 快讯
MarkTechPost 官方资讯

MarkTechPost:Anthropic Releases Claude Opus 5.5: Fable 5.1-Level Performance at 40% Lower Running Cost Th…

原文摘要:Anthropic has released Claude Opus 5.5, the first model in its new Claude 5.5 family. The team states it performs at the level of Claude Fable 5.1 on most work. It also costs 40% l 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:NVIDIA Introduces SoL-Pi: Auto-Research Loops That Cut Coding Agent Token Traffic by Up to 4…

原文摘要:NVIDIA researchers have released SoL-Pi, 4 harness mechanisms for the open-source Pi coding agent, discovered by an AI running auto-research loops across 535 environments. On EdgeB 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:SpaceXAI Releases Grok 4.7: A Larger Base Model at the Same $2/$6 Price as Grok 4.6

原文摘要:SpaceXAI has released Grok 4.7, its new flagship model for coding, agentic tasks, and knowledge work. Grok 4.7 is built on a larger base model and a longer reinforcement learning r 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 21 日 2026-09-21 快讯
MarkTechPost 官方资讯

MarkTechPost:AWS Strands Agents Team Releases Strands Harness: An Open-Source Agent Harness With 28% Lowe…

原文摘要:Many 开发者 find that an agent idea works inside Claude Code or Codex, then struggles once they rebuild it with their own loop. The Strands Agents team at AWS is targeting that 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:Alibaba Qwen Releases Qwen-Image-2.1: A 7B Open-Weight Model for Image Generation and Editin…

原文摘要:Alibaba's Qwen team has released Qwen-Image-2.1, a 7B diffusion transformer that handles text-to-image generation, multi-reference editing, and native RGBA transparency in one chec 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 20 日 2026-09-20 快讯
MarkTechPost 官方资讯

MarkTechPost:You too Google! Google Confirms Gemini Breached 3 Companies in AI Security Tests

原文摘要:Google says Gemini accessed 3 real companies in May by guessing a password and reusing credentials from a public 代码仓库. Irregular told 4 labs in late July. Google spoke on Sep 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:Alibaba Qwen Team Releases Qwen3.8-LiveTranslate: A Real-Time Interpretation Model That Cuts…

原文摘要:Qwen has released Qwen3.8-LiveTranslate, a real-time simultaneous interpretation model built on a new Interleave architecture. It cuts average lagging (LAAL) from 2.8 seconds to 2. 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 19 日 2026-09-19 快讯
MarkTechPost 官方资讯

MarkTechPost:TypeSafe AI Releases Jev: A System One Model That Returns Typed, Calibrated Decisions Instea…

原文摘要:TypeSafe AI released Jev, a System One model that answers typed questions with probabilities instead of generating text. Input costs $0.042 per 1M tokens, and output tokens are fre 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:Linkup Research Releases SPARSEUP: A 149M-Parameter Open-Source Sparse Embedding Model

原文摘要:Linkup Research has released SPARSEUP, an open-source sparse embedding model built on a 149M-parameter ModernBERT backbone. It scores 56.4 nDCG@10 on BEIR-13, which Linkup calls th 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:Meta Launches Muse for Mac: A Personal AI Agent That Works Across Your Files, Mail, Messages…

原文摘要:Meta has released Muse for Mac, the first version of Muse that can complete things on a user’s computer. The agent works with local files and native apps, where your data already l 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:SpaceXAI Releases Grok Voice Transcribe 2.0: A Speech-to-Text API Claiming 2x Accuracy Over …

原文摘要:SpaceXAI has released Grok Voice Transcribe 2.0, its newest speech-to-text model for batch and streaming audio. The company says it is twice as accurate as version 1.0 at the same 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 18 日 2026-09-18 快讯
MarkTechPost 官方资讯

MarkTechPost:Jina AI Releases jina-ocr-v1: A 3.4B MoE Document Parser With Built-In Speculative Decoding …

原文摘要:Jina AI has released jina-ocr-v1, a visual document parser that converts PDFs, scans, tables, charts and invoices into Markdown. The model has 3.4B total parameters, with about 570 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

Qwen3.8-Omni-Flash 发布:百万上下文、音视频理解与工具调用

阿里 Qwen 发布 Qwen3.8-Omni-Flash,这是一个支持 100 万上下文的 omni-modal 模型,围绕 Agentic 音视频理解与工具使用设计,官方称在 OmniVideoBench 上减少约 45.7% 的 token 消耗。它值得关注的点有两个:一是把音频、视频理解和任务规划、工具调用放进同一个模型,二是长上下文叠加 token 效率优化,直接关系到 Agent 处理长视频或多模态任务的成本。影响的主要是做视频理解、会议分析、多模态 Agent 的开发者,以及需要处理长音视频素材的内容团队。验证方式:通过 Qwen 官方渠道申请或调用该模型,用一段带语音的长视频测试摘要与工具调用,记录实际 token 用量和延迟,再与现有方案做同任务对比。

MarkTechPost 官方资讯

MarkTechPost:Salesforce Agentforce: Bridging the Enterprise AI Gap from ‘Vibe Coding’ to Battle-Tested Or…

原文摘要:Building an AI prototype is easy, but operating autonomous agents at scale requires production-grade tooling. Salesforce Agentforce bridges the gap from "vibe coding" to enterprise 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 17 日 2026-09-17 快讯
MarkTechPost 官方资讯

MarkTechPost:Microsoft Open-Sources TauGrid: A Kubernetes-Native Stack for GPU AI Workloads

原文摘要:Microsoft's AKS engineering team open-sourced TauGrid on August 28, 2026, packaging the tau CLI, Kueue queueing, KubeRay orchestration, GPU node health monitoring and observability 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:OpenAI Releases a Model Misalignment Disclosure Framework With 3 Review Tracks and 6 Inciden…

原文摘要:OpenAI can disclose misalignment before fixes exist. Its 6 initial reports include fabricated data and leaked API keys. The post OpenAI Releases a Model Misalignment Disclosure Fra 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:Google Research Introduces Retrieve-for-Train (R4T): An RL-Compiled Diffusion Retriever for …

原文摘要:Google Research has introduced Retrieve-for-Train (R4T), a framework for search that returns coherent, diverse result sets. It trains a fan-out language model with RL once, using g 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:Nunchux AI Introduces VC-Attention: A Training-Free Low-Bit Attention Kernel That Speeds Up …

原文摘要:Nunchux AI has released VC-Attention, a training-free low-bit attention kernel built for video Diffusion Transformers (DiTs). It targets 2 problems at once: value quantization erro 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 16 日 2026-09-16 快讯
MarkTechPost 官方资讯

MarkTechPost:Stanford Researchers Release Paper2Agent: Turning Research Papers Into AI Agents That Reprod…

原文摘要:Paper2Agent, published in Nature, converts papers into validated MCP tools, scoring 91.2% on 300 questions across 74 papers. The post Stanford Researchers Release Paper2Agent: Turn 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:Prior Labs Releases TabPFN-3.5: A Tabular Foundation Model That Beats the Winning Otto Kaggl…

原文摘要:Prior Labs released TabPFN-3.5, a tabular foundation model pretrained only on synthetic data that beats Otto's winning solution. The post Prior Labs Releases TabPFN-3.5: A Tabular 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:Nums AI Releases Causilo: A Tabular Foundation Model That Tops TabArena Among Single Models

原文摘要:Nums AI has released Causilo, a pretrained tabular foundation model for classification and regression with a scikit-learn interface. It posts the top TabArena Elo among single mode 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 15 日 2026-09-15 快讯
MarkTechPost 官方资讯

MarkTechPost:Inside NVIDIA’s cuDNN Graph API: Fusion, Autotuning, and Plan Reuse with cuDNN Frontend

原文摘要:Learn how to leverage NVIDIA’s cuDNN Frontend Graph API to build custom kernel fusions, autotuning engine configurations, FP8-style epilogues, scaled dot-product attention, dynamic 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:Google Releases Gemini 3.8 Live and 3.8 Live Extended Thinking for Production Grade Voice Ag…

原文摘要:Google has released Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, its most advanced live dialogue models to date. The models execute tools and API calls in the background 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:Agent-net Open Sources Webagent: A Go Harness That Turns Any Website into a Guarded AI Agent

原文摘要:Agent-net, the team building an agent-to-agent marketplace where AI agents discover, trust, and pay each other, has released Webagent, an 开源 harness for standing up public 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 14 日 2026-09-14 快讯
MarkTechPost 官方资讯

MarkTechPost:Agent Harness vs Agent Framework vs MCP: Which Layer Owns the Loop, State, Tools, Permission…

原文摘要:A practitioner's map of the 3 layers in a modern agent stack, with verified sources and an overlap analysis. The post Agent Harness vs Agent Framework vs MCP: Which Layer Owns the 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:Reward AI Releases OM-1: A Robot Policy Trained on Human Demonstrations Only, With No Teleop…

原文摘要:Reward AI has released OM-1 (Omnibody Model 1), a general-purpose manipulation policy trained entirely on human demonstrations captured with a 7-DoF wearable glove, with no teleope 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:Sakana AI Researchers Introduce PC-ALM, a Layer-Local Alternative to Backpropagation That Tr…

原文摘要:Sakana AI researchers Jeffrey Seely and Julian Gould introduce Augmented Lagrangian Predictive Coding (PC-ALM), a local-learning alternative to backpropagation. By attaching a Lagr 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:NVIDIA Open-Sources OSMO: One YAML Orchestrates Physical AI Training, Simulation, and Robot …

原文摘要:NVIDIA has open-sourced OSMO, the Kubernetes-native 工作流 orchestrator it uses internally for Project GR00T, Isaac Lab, and Isaac Sim. OSMO lets robotics teams define training, 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:Anthropic’s 3-Step ‘Pace the Frontier’ Plan Wins OpenAI, xAI and Microsoft Support: Is It To…

原文摘要:Dario Amodei published "We Must Pace the Frontier," and Sam Altman, Elon Musk and Satya Nadella endorsed it within a day. The trigger was a July incident in which roughly 1,200 Ope 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 13 日 2026-09-13 快讯
MarkTechPost 官方资讯

MarkTechPost:Hierarchical NeRF with JAX3D for Volumetric Rendering, Novel-View Synthesis, and 3D Reconstr…

原文摘要:In this tutorial, we build an end-to-end hierarchical Neural Radiance Field (NeRF) using JAX, Flax, Optax, and the volume-rendering primitives provided by jax3d. We first construct 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:A Princeton Researcher Proposes Recurrent Looped Transformer (RLT) that Carries Decoder Stat…

原文摘要:Yifan Zhang's Recurrent Looped Transformer (RLT) technical report proposes a causal encoder paired with a recurrent decoder that carries its final hidden state and layerwise slidin 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:Context Engineering Inside the Harness: 4 Mechanisms That Beat Context Overflow and Goal Los…

原文摘要:A shallow agent is an LLM calling tools in a loop, and on long tasks it fails in 2 ways: context overflow and goal loss. This article opens the harness layer that fixes both, with 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:Implementation of Machine Learning 工作流 with NVIDIA cuML, RAPIDS, GPU Benchmarking, Exp…

原文摘要:This practical tutorial demonstrates how to build and accelerate machine learning 工作流 using NVIDIA cuML and RAPIDS. It covers GPU environment setup, zero-code scikit-learn ac 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 12 日 2026-09-12 快讯
MarkTechPost 官方资讯

MarkTechPost:Cognition Releases SWE-2: A Kimi K3 Post-Trained Coding Model That Matches Fable 5.1 on Fron…

原文摘要:Cognition, the company behind the Devin coding agent, has released SWE-2, its most capable coding model to date. SWE-2 is post-trained with reinforcement learning from Kimi K3, Moo 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。