搜索:Research

共命中 50 条(服务端检索)
从搜索框到 Deep Research:Agentic 检索怎么把一次查询变成一场调查
Google 比 OpenAI 早七周发布 Deep Research,但把「研究计划让用户改批」产品化的也是它——agentic 检索的通用循环是:规划、迭代检索、阅读、反思补漏、交叉验证、带引用报告。本文拆解这个循环的两种实现路线(端到端 RL vs 显式编排),用 BrowseComp 上「裸模型不足 10% vs deep research 51.5%」的差距说明多步浏览行为本身值多少分,也把「慢不等于对」的引用可靠性研究摆上台面。
原创 智能体 精选 · 原创 · 今天 阅读 1·访客 1
Introducing Quine: An AI research system designed for the complexity of biology
At a glance Quine (opens in new tab) is a research effort to create a multimodal world model of biology and an interacti…
行业动态 Microsoft Research · 9-29 阅读 24·访客 24
Nous Research confirms it hit $1.5B valuation, launches AI agents for business users
Nous Research, the startup developing the open source Hermes Agent, has raised a $90 million Series B at a $1.5 billion …
智能体 TechCrunch · 今天 阅读 1·访客 1
A Coding Guide to Google Research’s Kauldron: Configs That Are Plain Data, Components Wired by String, and a JAX Trainer You Can Read End to End
In this tutorial, we implement **Kauldron**, the JAX training library from Google Research that describes itself as opti…
行业动态 MarkTechPost · 6天前 阅读 7·访客 7
Google Research Open-Sources RRSI: AI Agents That Improve Their Own Harness Without Overfitting
Google Cloud AI Research**, with UNC-Chapel Hill, Stanford and Washington University in St. Louis, has released RRSI (Re…
智能体 MarkTechPost · 9-29 阅读 25·访客 25
Google Research Introduces an AI Video Co-Director: 4 Agentic Frameworks for Coherent, Minutes-Long Video Generation
Google Research has introduced an **AI video co-director** for long-form video generation. The suite of 4 agentic framew…
智能体 MarkTechPost · 9-28 阅读 24·访客 24
A Coding Guide to Google Research’s MSEB: Writing Sound Encoders to the Benchmark Contract and Scoring Them Across Classification, Clustering, Retrieval and Segmentation
In this tutorial, we work with **MSEB**, the Massive Sound Embedding Benchmark from Google Research, and approach it fro…
研究前沿 MarkTechPost · 9-27 阅读 32·访客 31
OpenAI says 80 to 90 percent of its research already targets GPT 7 and beyond
Sep 27, 2026 Boris Power, OpenAI's Head of Applied Research, says 80 to 90 percent of the company's research goes toward…
大模型 The Decoder · 9-27 阅读 28·访客 28
Exa Launches Agent Ultra: A Subagent Swarm Deep Research API Built for Exhaustive List Building
Exa has released **Agent Ultra**, the highest effort level of its Exa Agent API. It is built for research that must run …
智能体 MarkTechPost · 9-26 阅读 28·访客 28
Inside Basecamp Research, the AI startup turning evolution into training data
Sep 23, 2026 Basecamp Research / GPT-Image-2 prompted by THE DECODER Basecamp Research has raised $140 million from inve…
大模型 The Decoder · 9-23 阅读 25·访客 25
Recursive self-improvement of AI research agents
AI agents are beginning to automate research and development across the AI stack, from improving training efficiency to …
智能体 HuggingFace Daily Papers · 9-22 阅读 11·访客 11
Linkup Research Releases SPARSEUP: A 149M-Parameter Open-Source Sparse Embedding Model
Linkup Research has released SPARSEUP, an open-source sparse embedding model built on a 149M-parameter ModernBERT backbo…
开源项目 MarkTechPost · 9-19 阅读 32·访客 29
Google Research Introduces Retrieve-for-Train (R4T): An RL-Compiled Diffusion Retriever for 12× to 20× Faster Query Fan-Out
Google Research has introduced Retrieve-for-Train (R4T), a framework for search that returns coherent, diverse result se…
行业动态 MarkTechPost · 9-17 阅读 20·访客 20
IdeaAMBIG: Benchmarking Implementation-Critical Gaps in Research-Idea Specifications
A research idea may be novel, coherent, and scientifically plausible, yet its proposed method may remain insufficiently …
研究前沿 HuggingFace Daily Papers · 9-9 阅读 16·访客 13
SAEScientist-Bench: Can AI Agents Conduct Autonomous SAE Interpretability Research?
While research on recursive self-improvement (RSI) has predominantly automated model training pipelines, reliable autono…
智能体 HuggingFace Daily Papers · 9-8 阅读 29·访客 28
Scores Alone Do Not Prove Discovery: The Discovery Certification Protocol for Auditing AI Research Agents
AI research agents combine prior knowledge, public sources, and experimental feedback to produce useful results. The Dis…
智能体 HuggingFace Daily Papers · 9-7 阅读 23·访客 21
Scaling Automatic Research Agents via World Models
Automating empirical research is a long-standing direction of AI. Recent automatic research (AutoResearch) agents bring …
智能体 HuggingFace Daily Papers · 8-29 阅读 19·访客 17
Agent Lightning v1.0: A 3,500-Line Lightweight Agentic RL Framework for Training Agents with Real Harnesses
At a glance Harnessed Agentic RL: Microsoft Research Asia introduces a training paradigm in which the same agent harness…
智能体 Microsoft Research · 今天 阅读 4·访客 4
What AI gets wrong and what failure teaches us
Jennifer Neville is a partner research manager at Microsoft who’s built a career around understanding and advancing AI f…
行业动态 Microsoft Research · 昨天 阅读 1·访客 1
Offloaded inference for real-world physical AI robotics
At a glance Challenges a core assumption in robotics AI: Our research shows that running physical AI inference exclusive…
智能体 Microsoft Research · 9-24 阅读 37·访客 37
Google Research Moves Federated Learning Into TEEs: Gboard Now Trains With Externally Verifiable Differential Privacy
!(https://www.gstatic.com/images/branding/googleg/1x/googleg_standard_color_128dp.png)Add as a preferredsource on Google…
行业动态 MarkTechPost · 4天前 阅读 18·访客 18
ScholarCatalyst: A Benchmark for Retrieving Papers That Inspire New Research
What makes great scientists great? Even as AI systems start to make progress on open problems, scientists remain far ahe…
研究前沿 HuggingFace Daily Papers · 10-1 阅读 8·访客 8
The Download: OpenAI’s chief research officer explains its hacking response
This is today's edition of* *The Download*,*our weekday newsletter that provides a daily dose of what's going on in the …
行业动态 MIT Technology Review · 9-30 阅读 14·访客 14
“We’re not going to shoot ourselves in the foot” over hack fallout, says OpenAI’s chief research officer
Two months after the bombshell news that a swarm of its agents had broken their containment and hacked into the computer…
智能体 MIT Technology Review · 9-30 阅读 7·访客 7
NVIDIA Introduces SoL-Pi: Auto-Research Loops That Cut Coding Agent Token Traffic by Up to 49%
Coding agents now run for hours, not minutes. Every edit, test run and log read goes back into the model’s context. A te…
智能体 MarkTechPost · 9-22 阅读 33·访客 32
Anthropic wants you to know Claude leads a quarter of its research, but "lead" doesn't mean what you think
For the first time, Anthropic is releasing metrics on how it builds its own AI. Claude already "leads" 26 percent of the…
大模型 The Decoder · 9-18 阅读 18·访客 17
Gricea: An Open Science Platform for Conversational AI Research
We need studies on conversational AI (CAI) at scale to understand human behavior and shape CAI design. However, fragment…
行业动态 HuggingFace Daily Papers · 9-18 阅读 16·访客 16
Stanford Researchers Release Paper2Agent: Turning Research Papers Into AI Agents That Reproduce Results and Run on New Data
Paper2Agent, published in Nature, converts papers into validated MCP tools, scoring 91.2% on 300 questions across 74 pap…
智能体 MarkTechPost · 9-17 阅读 32·访客 32
SoL-Pi: Recursively Scaling Auto-Research Loops for Efficient Agent Harness
As coding agents move from supervised code completion to unattended, around-the-clock exploration, their work expands fr…
智能体 HuggingFace Daily Papers · 9-17 阅读 26·访客 25
VidaForge: Open Research Infrastructure for Video Pretraining Data Recipes
Video foundation models increasingly rely on large-scale pretraining data, yet the end-to-end data pipelines behind them…
行业动态 HuggingFace Daily Papers · 9-6 阅读 15·访客 13
Dr. Claw: An AI Scientist Workspace for Vibe Research
Command-line coding agents (e.g., Claude Code, Gemini CLI) can already read and write files and sustain long sessions, y…
智能体 HuggingFace Daily Papers · 8-31 阅读 19·访客 17
Perplexity Releases pplx-embed-v2-context-9b-preview: A Contextual Embedding Model That Retrieves Answers and Their Supporting Evidence
Perplexity Research and turbopuffer have released **pplx-embed-v2-context-9b-preview**, a contextual embedding model for…
行业动态 MarkTechPost · 10-1 阅读 14·访客 14
Fireworks AI Releases Ember-1: A Post-Trained Kimi K3 That Uses About 40% Fewer Tokens
Fireworks AI has released Ember-1, a specialized model from Fireworks Research built by post-training Moonshot AI’s open…
研究前沿 MarkTechPost · 9-28 阅读 35·访客 35
AI agents do more of the work in model development, but humans still make the decisions
Sep 27, 2026 Nano Banana Pro prompted by THE DECODER A research team documented how humans and AI agents worked together…
智能体 The Decoder · 9-27 阅读 27·访客 27
Counterpoint:2025-2030 年间,全球 200 美元以下智能手机年出货将减少超 2.3 亿部
IT之家 9 月 25 日消息,当地时间 24 日,Counterpoint Research 发布的最新《按价格区间划分的智能手机出货量预测》追踪报告预计,2025 年至 2030 年间,全球 200 美元(IT之家注:现汇率约合 1,3…
行业动态 IT之家 · 9-25 阅读 12·访客 12
Perplexity Trains Its Computer Agent on Real Mistakes With Hint-Guided Self-Distillation
Perplexity Research published a new post-training study. It trains a model inside Perplexity Computer on real user sessi…
智能体 MarkTechPost · 9-25 阅读 22·访客 21
How Open Science Can Help Researchers Prepare for the Next Pandemic
When COVID-19 emerged, scientists had a crucial advantage: Decades of prior research on coronaviruses meant they underst…
行业动态 NVIDIA Blog · 9-24 阅读 19·访客 19
WhatWorkedBench: Benchmarking Experimental Understanding in AI Agents
AI research agents need reliable knowledge of how their experiments change outcomes. We introduce WhatWorkedBench to mea…
智能体 HuggingFace Daily Papers · 9-23 阅读 10·访客 10
Nokia Open-Sources AnyJev: A Training-Free Layer That Turns Any Open LLM Into a Calibrated Decision Model
Nokia’s applied research team has open-sourced AnyJev, a Python library that turns an open LLM into a decision model. It…
开源项目 MarkTechPost · 9-23 阅读 70·访客 68
Why Deploying Physical AI at Scale Demands Safety at Every Layer
Physical AI is moving rapidly from research to large-scale deployment. By 2035, ABI Research projects an installed base …
行业动态 NVIDIA Blog · 9-22 阅读 33·访客 33
2026 上半年全球 AI 眼镜出货量同比大增 263%,“无屏幕”产品占主流
IT之家 9 月 22 日消息,外媒 Android Authority 援引 Counterpoint Research《智能眼镜市场追踪“报告,认为目前智能眼镜仍属于相对小众的科技产品,但其市场增长速度正在明显加快。 数据显示, 202…
行业动态 IT之家 · 9-22 阅读 31·访客 31
OpenAI forms math advisory group as its AI resolves more than 100 open problems
The group won't be given leeway to slow down or redirect OpenAI's ongoing mathematical research.]
行业动态 TechCrunch · 9-22 阅读 24·访客 24
The US Navy just told us what’s on its tech wish list for the next several years
Navy CTO Justin Fanelli talks co-investing alongside VCs instead of funding early research himself, recent buys like a $…
行业动态 TechCrunch · 9-20 阅读 31·访客 31
AI hallucination nearly triggers US military operation
“It’s important for service members to understand the uncertainty inherent to LLMs," a GovAI research scholar warns. ]
大模型 TechCrunch · 9-19 阅读 25·访客 25
Base Labs launches an open-weight AI safety partnership with Hugging Face and Goodfire
Base Labs, the research group Baseten spun up earlier this year, will develop and publish methods for training and monit…
行业动态 TechCrunch · 9-18 阅读 17·访客 17
Google Deepmind launches interdisciplinary institute to tackle the big questions around AGI
Google Deepmind has founded the Deepmind Institute (DMI), an interdisciplinary research platform focused on AGI. Led by …
行业动态 The Decoder · 9-17 阅读 14·访客 14
Agora: Git as Shared Memory for Collective AutoResearch
Autonomous research loops such as AutoResearch show that one coding agent can improve a training setup unattended. Run s…
智能体 HuggingFace Daily Papers · 9-16 阅读 20·访客 19
ALPINE: Adaptive Localization for Parameter- and Sample-Efficient Few-Shot Learning
Few-shot learning research is predominantly evaluated on accuracy alone, with limited attention to the parameter and tra…
行业动态 HuggingFace Daily Papers · 9-16 阅读 8·访客 8
ScienceBuddy: Recursive-in-Recursive Self-Improvement for Interactive Scientific Agents
We introduce and release ScienceBuddy, an interactive scientific research workspace that brings continually improving sc…
智能体 HuggingFace Daily Papers · 9-15 阅读 18·访客 17
Zuckerberg's Biohub leads a $1.8 billion push to build AI models that predict cell behavior
Manuel Uth Oct 7, 2026 AI models are supposed to learn to predict cell behavior, which could speed up drug development.*…
行业动态 The Decoder · 今天 阅读 4·访客 4