AI 每日快讯

AI 每日快讯

AI 产品、模型、开源工具和官方动态的时间流。保留历史记录,按分类、日期和标签继续筛选。

3779历史快讯
207开源工具
80当前结果
09 月 30 日 2026-09-30 快讯
MarkTechPost 官方资讯

MarkTechPost:OpenAI Releases GPT-6.1 Sol: Near-Astra Coding and Computer Use at One-Fifth of Astra’s Toke…

原文摘要:OpenAI released GPT-6.1 Sol on September 29, 2026, an upgrade to GPT-6 Sol. It reaches near-Astra results on agentic coding, computer use and professional work at one-fifth of Astr 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:One Bad Prompt Took Down a Company’s Salesforce: RSA’s Jim Taylor on Agent ID and Taming the…

原文摘要:RSA launched Agent ID at The AI Conference in San Francisco. It's an agentic identity security platform for finance, government, healthcare, and critical infrastructure. It has 3 m 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 29 日 2026-09-29 快讯
MarkTechPost 官方资讯

MarkTechPost:OpenAI Launches dots: Always-On GPT-6 Astra Agents That Work From Their Own Cloud Computers

原文摘要:OpenAI just introduced dots at their DevDay today. Dots are persistent AI agents powered by GPT-6 Astra. Each dot gets its own cloud computer and browser. It works across 4,000+ ap 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:Google Research Open-Sources RRSI: AI Agents That Improve Their Own Harness Without Overfitt…

原文摘要:Google Cloud AI Research has open-sourced RRSI, a framework that lets LLM agents rewrite their own prompts, tools and memory while model weights stay frozen. It adds a leakage crit 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:H Company Releases Holo4: Open-Weight Computer-Use Models That Click, Code and Call Tools Ac…

原文摘要:H Company has released Holo4, a family of generalist computer-use models for AI agents. One set of weights clicks and types on screens. It also writes code and calls MCP or API too 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 28 日 2026-09-28 快讯
MarkTechPost 官方资讯

MarkTechPost:NVIDIA Launches Open Agent Safety Platform: OpenShell Sandboxes Agents on Vera CPUs While Se…

原文摘要:NVIDIA has launched the Open Agent Safety Platform, an open reference design that enforces AI agent safety outside the agent itself. OpenShell, an Apache 2.0 runtime, sandboxes age 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 26 日 2026-09-26 快讯
MarkTechPost 官方资讯

MarkTechPost:Exa Launches Agent Ultra: A Subagent Swarm Deep Research API Built for Exhaustive List Build…

原文摘要:Exa has released Agent Ultra, the highest effort mode of its Exa Agent API. It coordinates subagents across thousands of sources for list building and entity enrichment. Exa report 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 25 日 2026-09-25 快讯
MarkTechPost 官方资讯

MarkTechPost:Perplexity Trains Its Computer Agent on Real Mistakes With Hint-Guided Self-Distillation

原文摘要:Perplexity Research published a new post-training study. It trains a model inside Perplexity Computer on real user sessions, including failed ones. The method pairs rejection sampl 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:Fastino Releases GLiNER2.5-Decide: A 340M Open-Weight Decision Model That Runs on CPU

原文摘要:Fastino Labs has released GLiNER2.5-Decide, a 340M-parameter open-weight decision model. It takes text and a schema of typed questions and returns structured answers. Each answer c 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 24 日 2026-09-24 快讯
MarkTechPost 官方资讯

MarkTechPost:Contrastive-LM Releases CLM-8B: An Open System One Model That Scores Agent Actions Up to 9× …

原文摘要:Contrastive-LM has released CLM-8B, an open System One model that scores candidate actions against a state instead of generating text. It adds 2 small projection heads to a frozen 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:A Coding Guide to TypeSafe AI Jev: Typed Decisions, Calibrated Confidence, and Speculative F…

原文摘要:This tutorial provides a complete coding guide to TypeSafe AI's Jev, a System One model designed for non-text, structured judgments. It covers installing the official Python SDK, u 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 23 日 2026-09-23 快讯
MarkTechPost 官方资讯

MarkTechPost:Google Releases Gemini 3.8 Flash TTS and Flash-Lite TTS With Prompt-Based Voice Design

原文摘要:Google has released Gemini 3.8 Flash TTS and Flash-Lite TTS, 2 new text-to-speech models available now through the Gemini API and Google AI Studio. Flash TTS designs new voices fro 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 22 日 2026-09-22 快讯
MarkTechPost 官方资讯

MarkTechPost:Anthropic Releases Claude Opus 5.5: Fable 5.1-Level Performance at 40% Lower Running Cost Th…

原文摘要:Anthropic has released Claude Opus 5.5, the first model in its new Claude 5.5 family. The team states it performs at the level of Claude Fable 5.1 on most work. It also costs 40% l 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:NVIDIA Introduces SoL-Pi: Auto-Research Loops That Cut Coding Agent Token Traffic by Up to 4…

原文摘要:NVIDIA researchers have released SoL-Pi, 4 harness mechanisms for the open-source Pi coding agent, discovered by an AI running auto-research loops across 535 environments. On EdgeB 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:SpaceXAI Releases Grok 4.7: A Larger Base Model at the Same $2/$6 Price as Grok 4.6

原文摘要:SpaceXAI has released Grok 4.7, its new flagship model for coding, agentic tasks, and knowledge work. Grok 4.7 is built on a larger base model and a longer reinforcement learning r 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 21 日 2026-09-21 快讯
MarkTechPost 官方资讯

MarkTechPost:AWS Strands Agents Team Releases Strands Harness: An Open-Source Agent Harness With 28% Lowe…

原文摘要:Many 开发者 find that an agent idea works inside Claude Code or Codex, then struggles once they rebuild it with their own loop. The Strands Agents team at AWS is targeting that 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 19 日 2026-09-19 快讯
MarkTechPost 官方资讯

MarkTechPost:Meta Launches Muse for Mac: A Personal AI Agent That Works Across Your Files, Mail, Messages…

原文摘要:Meta has released Muse for Mac, the first version of Muse that can complete things on a user’s computer. The agent works with local files and native apps, where your data already l 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 18 日 2026-09-18 快讯
MarkTechPost 官方资讯

MarkTechPost:Salesforce Agentforce: Bridging the Enterprise AI Gap from ‘Vibe Coding’ to Battle-Tested Or…

原文摘要:Building an AI prototype is easy, but operating autonomous agents at scale requires production-grade tooling. Salesforce Agentforce bridges the gap from "vibe coding" to enterprise 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 16 日 2026-09-16 快讯
MarkTechPost 官方资讯

MarkTechPost:Stanford Researchers Release Paper2Agent: Turning Research Papers Into AI Agents That Reprod…

原文摘要:Paper2Agent, published in Nature, converts papers into validated MCP tools, scoring 91.2% on 300 questions across 74 papers. The post Stanford Researchers Release Paper2Agent: Turn 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 15 日 2026-09-15 快讯
MarkTechPost 官方资讯

MarkTechPost:Google Releases Gemini 3.8 Live and 3.8 Live Extended Thinking for Production Grade Voice Ag…

原文摘要:Google has released Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, its most advanced live dialogue models to date. The models execute tools and API calls in the background 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:Agent-net Open Sources Webagent: A Go Harness That Turns Any Website into a Guarded AI Agent

原文摘要:Agent-net, the team building an agent-to-agent marketplace where AI agents discover, trust, and pay each other, has released Webagent, an 开源 harness for standing up public 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 14 日 2026-09-14 快讯
MarkTechPost 官方资讯

MarkTechPost:Agent Harness vs Agent Framework vs MCP: Which Layer Owns the Loop, State, Tools, Permission…

原文摘要:A practitioner's map of the 3 layers in a modern agent stack, with verified sources and an overlap analysis. The post Agent Harness vs Agent Framework vs MCP: Which Layer Owns the 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:NVIDIA Open-Sources OSMO: One YAML Orchestrates Physical AI Training, Simulation, and Robot …

原文摘要:NVIDIA has open-sourced OSMO, the Kubernetes-native 工作流 orchestrator it uses internally for Project GR00T, Isaac Lab, and Isaac Sim. OSMO lets robotics teams define training, 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:Anthropic’s 3-Step ‘Pace the Frontier’ Plan Wins OpenAI, xAI and Microsoft Support: Is It To…

原文摘要:Dario Amodei published "We Must Pace the Frontier," and Sam Altman, Elon Musk and Satya Nadella endorsed it within a day. The trigger was a July incident in which roughly 1,200 Ope 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 13 日 2026-09-13 快讯
MarkTechPost 官方资讯

MarkTechPost:Context Engineering Inside the Harness: 4 Mechanisms That Beat Context Overflow and Goal Los…

原文摘要:A shallow agent is an LLM calling tools in a loop, and on long tasks it fails in 2 ways: context overflow and goal loss. This article opens the harness layer that fixes both, with 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:Implementation of Machine Learning 工作流 with NVIDIA cuML, RAPIDS, GPU Benchmarking, Exp…

原文摘要:This practical tutorial demonstrates how to build and accelerate machine learning 工作流 using NVIDIA cuML and RAPIDS. It covers GPU environment setup, zero-code scikit-learn ac 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 12 日 2026-09-12 快讯
MarkTechPost 官方资讯

MarkTechPost:Cognition Releases SWE-2: A Kimi K3 Post-Trained Coding Model That Matches Fable 5.1 on Fron…

原文摘要:Cognition, the company behind the Devin coding agent, has released SWE-2, its most capable coding model to date. SWE-2 is post-trained with reinforcement learning from Kimi K3, Moo 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 11 日 2026-09-11 快讯
MarkTechPost 官方资讯

MarkTechPost:Can LLMs Engineer Their Own Agent Harness? ByteDance Seed’s HarnessDev Says Only 34 of 64 Ch…

原文摘要:ByteDance Seed, SUTD, Georgia Tech, M-A-P, and TokenWave.AI introduce HarnessDev, a 评测 that scores the runnable harness a model builds rather than the answer it returns. Sta 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:Anthropic Adds Plugin Evals to Claude Code: 6 Grader Types, a No-Plugin Baseline, and a CI G…

原文摘要:Anthropic has published a new plugin evals 工作流 for Claude Code. The claude plugin eval command runs a plugin against realistic prompts, grades what Claude produced, and compar 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:Sakana AI Launches Fugu Max and Fugu Ultra v2 for Cheaper, Stronger Multi-Agent Orchestratio…

原文摘要:Sakana AI has released Fugu Max and Fugu Ultra v2, 2 models built on the same learned orchestration architecture. Fugu Max routes tasks to lean open and specialized models, includi 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 10 日 2026-09-10 快讯
MarkTechPost 官方资讯

MarkTechPost:OpenAI Launches the Agents API in Public Beta, Putting the Codex Harness Behind One API Call

原文摘要:OpenAI has released the Agents API in public beta. It gives 开发者 the same harness and infrastructure that run Codex. OpenAI hosts and maintains the harness. 开发者 run th 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:LandingAI Releases Agentic Document Extraction Gen2 with DPT-3 Pro and DPT-3 Verity

原文摘要:LandingAI has shipped Agentic Document Extraction Gen2, a rebuild of its document stack on the DPT-3 model family. Chunks are retired in favor of a document, page and block tree. D 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 09 日 2026-09-09 快讯
MarkTechPost 官方资讯

MarkTechPost:Google Open-Sources Mantis: A Modular Skills Toolkit That Lets Coding Agents Find, Reproduce…

原文摘要:Google has open-sourced Mantis, a stack-agnostic toolkit of security review skills for AI coding agents. It runs the full vulnerability lifecycle: sweep the code, filter false posi 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:Gradium Launches Voice Design: Write a Prompt, Get a Brand New Synthetic Voice in Seconds

原文摘要:Voice agent teams keep hitting the same wall. The catalog holds 400 voices and the brief asks for the one that is not in it: a Quebecoise receptionist for a Montreal dealership, a 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:Meta Introduces Muse, a Personal AI Agent That Runs on Its Own Dedicated Secure Cloud Comput…

原文摘要:Today, Meta has introduced Muse, a personal AI agent that takes actions rather than just answering questions. Muse can send emails, book travel, negotiate bills, and pursue long te 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 08 日 2026-09-08 快讯
MarkTechPost 官方资讯

MarkTechPost:Reducto Releases r-1: A Single Pass Document Parsing Model That Cuts Errors 20% at 1 Cent Pe…

原文摘要:We look at r-1, the document parsing model Reducto released on September 1, 2026. We walk through how it folds OCR, layout detection, tables, formatting and grounding into one full 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 06 日 2026-09-06 快讯
MarkTechPost 官方资讯

MarkTechPost:Meta FAIR Introduces AI Research Preference Models (RPMs): Ranking ML Experiments Before Spe…

原文摘要:AI research agents can propose far more experiments than they can afford to run. Meta FAIR, Oxford and UCL introduce AI Research Preference Models — frozen LLM judges that rank 15 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:UC Berkeley Researchers Release CUA-Lite, an Open Platform Unifying Sandboxes, Data, Evaluat…

原文摘要:Training and benchmarking a computer-use agent needs four things — agents, environments, traces, and a framework to evaluate and train them — and all four ship in incompatible form 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 05 日 2026-09-05 快讯
MarkTechPost 官方资讯

MarkTechPost:GitHub Introduces Project HydraFusion: Runtime Multi-Model Orchestration That Builds a Workf…

原文摘要:We look at Project HydraFusion, GitHub's research preview that treats 工作流 selection as an optimization problem rather than a model picker. We break down the three execution pa 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:Google Launches Agentic Video Understanding for Gemini Flash Models, Cutting Video Tokens by…

原文摘要:Gemini now navigates video instead of ingesting it at 1 FPS, loading only the segments a prompt needs. The post Google Launches Agentic Video Understanding for Gemini Flash Models, 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:NVIDIA Releases Personal AI Router (PAIR): An 开源 Virtual Inference Router that Dist…

原文摘要:We look at NVIDIA Personal AI Router (PAIR), an 开源 virtual inference router that spreads local AI requests across the machines already on a home network. We cover how PAIR 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 03 日 2026-09-03 快讯
MarkTechPost 官方资讯

MarkTechPost:Anthropic Released Claude Commerce Agents: An Apache-2.0 Blueprint for Shopping and Merchant…

原文摘要:Most teams building a shopping assistant or agent rebuild the same scaffolding: an agent loop, a tool layer over the catalog, an approval gate, and an eval suite. Anthropic has now 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:Meta AI Released Muse Spark 1.3: An Agentic Coding Model That Uses ~20% Fewer Tool Calls and…

原文摘要:Perplexity has shipped hybrid compute for its Mac app, splitting a single Perplexity Computer task between frontier models in the cloud and a compact model running on the user's ma 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 02 日 2026-09-02 快讯
MarkTechPost 官方资讯

MarkTechPost:Perplexity Releases Hybrid Compute on Mac: Cloud Agents Orchestrate Down to a Local Model, G…

原文摘要:Agentic assistants have a structural problem: the context that makes them useful — deal documents, privileged files, client records — is exactly the context users cannot send to a 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

09 月 01 日 2026-09-01 快讯
MarkTechPost 官方资讯

MarkTechPost:Researchers from Princeton, Ant Group and Stanford Introduce AQuA: A Two-Part Agentic Framew…

原文摘要:Quantitative research agents that write their own experiments can corrupt the evidence they later learn from. A leaky feature that scores well gets stored as a successful precedent 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

08 月 31 日 2026-08-31 快讯
MarkTechPost 官方资讯

MarkTechPost:Keenable AI Open-Sources NEEDLE: A Live Search 评测 That Rebuilds Its Query Set Every H…

原文摘要:How do you 评测 a web search API when the thing being tested can read the answer key? A search agent has a fetch tool. If the gold labels sit in a public dataset, the agent ca 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

08 月 30 日 2026-08-30 快讯
MarkTechPost 官方资讯

MarkTechPost:Lowest-Latency Inference APIs for Voice and Realtime Agents: A Time to First Token TTFT-Firs…

原文摘要:Voice agents fail on latency long before they fail on intelligence. Time to first token is the metric most teams use to choose an inference API, and it is the right starting point 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:Google AI Introduces EnvHarness: A Programmable Layer That Turns Static Agent Environments I…

原文摘要:Google Cloud AI Research, with Washington University in St. Louis and UNC Chapel Hill, has released EnvHarness, an Apache-2.0 layer that turns a static agent 评测 into one tha 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:Anthropic Opens a Research Preview of the Model Hardware Standard (MHS): A Shared Specificat…

原文摘要:Anthropic has opened a research preview of the Model Hardware Standard (MHS), a shared driver specification that lets AI agents discover and safely operate physical devices. Instru 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

08 月 29 日 2026-08-29 快讯
MarkTechPost 官方资讯

MarkTechPost:Building Custom Batched Ensemble Weather Forecasting with NVIDIA Earth2Studio

原文摘要:In this tutorial, we build an ensemble weather forecasting 工作流 with NVIDIA Earth2Studio. We install the required Earth2Studio components while preserving Colab’s existing CUDA 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

08 月 28 日 2026-08-28 快讯
MarkTechPost 官方资讯

MarkTechPost:Google AI Releases Gemini 3.5 Transcribe: A Speech-to-Text Model Reporting 2.6% Average WER …

原文摘要:Google has released Gemini 3.5 Transcribe, a speech-to-text model that ships as two separate endpoints rather than one. The streaming endpoint delivers sub-second transcription but 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

08 月 27 日 2026-08-27 快讯
MarkTechPost 官方资讯

MarkTechPost:Best Agent Sandboxes in 2026: Cold Start, Per-Second Pricing, and Network Policy Across E2B,…

原文摘要:Every agent that writes code needs somewhere to run it, and no two vendors quote the same units. This comparison measures burst cold start across E2B, Daytona, Modal, Cloudflare, a 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

08 月 26 日 2026-08-26 快讯
MarkTechPost 官方资讯

MarkTechPost:IBM Releases Granite 4.2: Bringing Native Reasoning and Agentic RL to Open Enterprise Models

原文摘要:IBM has released Granite 4.2, a family of open reasoning language models in 3B, 8B, and 30B sizes, all under Apache 2.0. Every model exposes a thinking / low-effort / non-thinking 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

08 月 24 日 2026-08-24 快讯
MarkTechPost 官方资讯

MarkTechPost:Scientific Data Analysis with LabPlot in Python: Signal Processing, Spectral Peak Fitting, V…

原文摘要:In this tutorial, we explore a LabPlot-inspired scientific data analysis 工作流 in Python while preserving the structure and terminology of LabPlot’s aspect tree, analysis kernel 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

08 月 23 日 2026-08-23 快讯
MarkTechPost 官方资讯

MarkTechPost:Vercel Introduces ‘Is Agentic’, a Free Agent-Readiness Scoring Tool That Audits Public Websi…

原文摘要:Vercel and Ora launched Is Agentic, a free audit scoring website readiness for AI agents across 118 checks. The post Vercel Introduces ‘Is Agentic’, a Free Agent-Readiness Scoring 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

08 月 22 日 2026-08-22 快讯
MarkTechPost 官方资讯

MarkTechPost:Decoding AI’s Open-Source Course Maps Three Ways to Run an Agent Loop and the Provider Econo…

原文摘要:Most teams treat ‘which model’ as the important decision. The harness engineering literature keeps pointing somewhere else. In LangChain’s Terminal-Bench experiment, changing only 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

08 月 21 日 2026-08-21 快讯
MarkTechPost 官方资讯

MarkTechPost:Building Agentic Document Intelligence Pipelines: Creating Scientific Figures with AutoFigur…

原文摘要:This tutorial explores AutoFigure, a practical toolkit for generating professional scientific figures directly from text descriptions and research papers. We walk through setting u 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

08 月 20 日 2026-08-20 快讯
MarkTechPost 官方资讯

MarkTechPost:Auditing Preference Biases and Fine-Tuning Language Models with Direct Preference Optimizati…

原文摘要:This tutorial provides an end-to-end 工作流 for fine-tuning language models using Direct Preference Optimization (DPO). We demonstrate how to audit the Anthropic HH-RLHF dataset 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

08 月 18 日 2026-08-18 快讯
MarkTechPost 官方资讯

MarkTechPost:Meet SAM (Sovereign Agent Mesh): A Zero-Config, Zero-Trust P2P Network for AI Agents

原文摘要:Google has open-sourced SAM (Sovereign Agent Mesh) under Apache-2.0 — and it has nothing to do with Segment Anything. SAM is a zero-config, zero-trust P2P overlay that lets autonom 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:Nous Research Ships Bot Mode for Hermes Agent, Turning Agent Profiles Into a Roster of Named…

原文摘要:Nous Research has shipped Bot Mode for Hermes Agent, its MIT-licensed 开源 agent. Bot Mode replaces the single-agent session list with a roster of named bots. Each bot is a 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:ByteDance Seed and Tsinghua AIR Introduces CUDA Agent: A Large-Scale Agentic RL System for C…

原文摘要:ByteDance Seed and Tsinghua AIR have released CUDA Agent, an agentic reinforcement learning system that trains a large language model to write GPU kernels that beat a compiler. The 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

08 月 17 日 2026-08-17 快讯
MarkTechPost 官方资讯

MarkTechPost:DeepSeek AI Releases DeepSeek Harness in 开发者 Preview: An MIT-Licensed Agent Harness Wh…

原文摘要:DeepSeek Harness v0.1 is an MIT-licensed agent harness where every capability is a Cordis plugin. Four runtime modes, append-only session logs, and provider-agnostic model routing. 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

08 月 14 日 2026-08-14 快讯
MarkTechPost 官方资讯

MarkTechPost:Create a Reasoning-Focused LLM: A Practical Guide to Streaming, Curating, and Fine-Tuning th…

原文摘要:This tutorial provides a complete 工作流 for building a compact, reasoning-focused language model. By streaming the SupraLabs reasoning corpus from Hugging Face, we apply quality 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

08 月 13 日 2026-08-13 快讯
MarkTechPost 官方资讯

MarkTechPost:SpaceXAI Releases Grok 4.6: A 500K-Context Frontier Model Tuned for Long-Running Agents, Cod…

原文摘要:SpaceXAI released Grok 4.6 on August 12, 2026 — a post-training upgrade over Grok 4.5, not a larger base model. It ties GPT-5.6 Sol Max at 61 on the Artificial Analysis Intelligenc 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

08 月 12 日 2026-08-12 快讯
MarkTechPost 官方资讯

MarkTechPost:NVIDIA AI Releases Nemotron 3.5 Lightning: A 30B Open MoE with 3B Active Parameters, and NeM…

原文摘要:NVIDIA's open 30B MoE targets the agent execution layer, with Switchyard routing each step to the cheapest capable model. The post NVIDIA AI Releases Nemotron 3.5 Lightning: A 30B 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

08 月 10 日 2026-08-10 快讯
MarkTechPost 官方资讯

MarkTechPost:Meta AI Releases Muse Glimmer: A 30B Open-Weights Agentic Model That Runs on One Consumer GP…

原文摘要:Meta's Muse Glimmer is a 30B open-weights agentic model under Apache 2.0. It fits 24 GB VRAM and decodes 3.1x faster with DFlash speculation. The post Meta AI Releases Muse Glimmer 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

08 月 09 日 2026-08-09 快讯
MarkTechPost 官方资讯

MarkTechPost:IMDb Sentiment Analysis with DistilBERT LoRA, TF-IDF Baselines, Calibration, Interpretabilit…

原文摘要:This tutorial provides a comprehensive guide to building a robust sentiment analysis 工作流. By combining classical TF-IDF baselines with modern parameter-efficient fine-tuning ( 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

08 月 08 日 2026-08-08 快讯
MarkTechPost 官方资讯

MarkTechPost:Meet Shepherd: An Open-Source Python Substrate That Lets Meta-Agents Fork, Replay, and Rever…

原文摘要:Long agent runs accumulate state that no transcript records — edited files, a live dev server, installed packages, a warm prompt cache. When an agent misreads a traceback at step 1 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

MarkTechPost 官方资讯

MarkTechPost:Pokee AI Releases Pokee-Isaac 28B: A 10M-Token Context Agentic Model Built to Run Inside the…

原文摘要:Pokee AI released Pokee-Isaac 28B, a 28B text-only foundation model with a 10M-token context window built to run inside the customer boundary. It scores 93.3% on RULER at 10M token 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。

08 月 07 日 2026-08-07 快讯
MarkTechPost 官方资讯

MarkTechPost:NVIDIA AI Releases NOOA: An Object-Oriented Python Framework That Turns an AI Agent Into a S…

原文摘要:NVIDIA Labs has open-sourced NOOA (NVIDIA Object-Oriented Agents), a model-agnostic Python framework for building AI agents. Agent development today is split across prompt template 来源:MarkTechPost。建议继续查看原文,重点核对它影响的工具入口、成本、风险和真实使用场景。