AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,973 results
15 Apr 2026

AgenticAI-DialogGen: Topic-Guided Conversation Generation for Fine-Tuning and Evaluating Short- and Long-Term Memories of LLMs

AgentsDGX agent

arXiv:2604.12179v1 Announce Type: new Abstract: Recent advancements in Large Language Models (LLMs) have improved their ability to process extended conversational contexts, yet fine-tuning and evaluat

Bad data, not bad AI, is what’s stalling enterprise deployments

AgentsDGX agent

The question is no longer whether to deploy AI — it’s why so many deployments stall before delivering returns. The answer usually comes down to a lack of trusted data foundation. As research from Qlik

Build once. Reuse everywhere. The Connectors API is now in Public Preview. Register MCP connectors once and use them across Le Chat, AI Stud…

AgentsDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Mistral AI has launched its Connectors API into Public Preview, enabling developers to register Model Context Protocol (MCP) connectors once and deploy them across multiple Mistral products. This 'bui

Instructions are all you need: Self-supervised Reinforcement Learning for Instruction Following

AgentsDGX agent

arXiv:2510.14420v4 Announce Type: replace-cross Abstract: Language models often struggle to follow multi-constraint instructions that are crucial for real-world applications. Existing reinforcement le

OVAL: Open-Vocabulary Augmented Memory Model for Lifelong Object Goal Navigation

AgentsDGX agent

arXiv:2604.12872v1 Announce Type: new Abstract: Object Goal Navigation (ObjectNav) refers to an agent navigating to an object in an unseen environment, which is an ability often required in the accomp

14 Apr 2026

A Benchmark and Multi-Agent System for Instruction-driven Cinematic Video Compilation

Model ReleasesDGX agent

arXiv:2604.10456v1 Announce Type: new Abstract: The surging demand for adapting long-form cinematic content into short videos has motivated the need for versatile automatic video compilation systems.

Big news for Indian customers 🇮🇳 You can now pay for Replit with UPI via @Razorpay, alongside debit & credit cards. And when you're ready …

AgentsDGX agent

Big news for Indian customers 🇮🇳 You can now pay for Replit with UPI via @Razorpay, alongside debit & credit cards. And when you're ready to monetize the app you built? Use Replit's Razorpay MCP to st

Both Ends Count! Just How Good are LLM Agents at 'Text-to-Big SQL'?

Model ReleasesDGX agent

arXiv:2602.21480v4 Announce Type: replace-cross Abstract: Text-to-SQL and Big Data are both extensively benchmarked fields, yet there is limited research that evaluates them jointly. In the real world

CARE-ECG: Causal Agent-based Reasoning for Explainable and Counterfactual ECG Interpretation

Model ReleasesDGX agent

arXiv:2604.10420v1 Announce Type: new Abstract: Large language models (LLMs) enable waveform-to-text ECG interpretation and interactive clinical questioning, yet most ECG-LLM systems still rely on wea

DeepReviewer 2.0: A Traceable Agentic System for Auditable Scientific Peer Review

Model ReleasesDGX agent

arXiv:2604.09590v1 Announce Type: new Abstract: Automated peer review is often framed as generating fluent critique, yet reviewers and area chairs need judgments they can audit: where a concern applie

EmbodiedGovBench: A Benchmark for Governance, Recovery, and Upgrade Safety in Embodied Agent Systems

Model ReleasesDGX agent

arXiv:2604.11174v1 Announce Type: cross Abstract: Recent progress in embodied AI has produced a growing ecosystem of robot policies, foundation models, and modular runtimes. However, current evaluatio

Escaping the Context Bottleneck: Active Context Curation for LLM Agents via Reinforcement Learning

Model ReleasesDGX agent

arXiv:2604.11462v1 Announce Type: new Abstract: Large Language Models (LLMs) struggle with long-horizon tasks due to the 'context bottleneck' and the 'lost-in-the-middle' phenomenon, where accumulated

From Understanding to Creation: A Prerequisite-Free AI Literacy Course with Technical Depth Across Majors

AgentsDGX agent

arXiv:2604.09634v1 Announce Type: cross Abstract: Most AI literacy courses for non-technical undergraduates emphasize conceptual breadth over technical depth. This paper describes UNIV 182, a prerequi

here’s a good application of harness permissions Programmatically Enforced Auto-Research: - auto-research loops usually expose a set of file…

Model ReleasesDGX agent

here’s a good application of harness permissions Programmatically Enforced Auto-Research: - auto-research loops usually expose a set of files that the agent is allowed to edit to hill climb a metric/e

HubSpot targets AI-driven buyer behavior shift with new tools and agents

IndustryDGX agent

Customer relationship management vendor HubSpot Inc. today introduced a set of product updates aimed at helping sales and marketing teams adapt to a shift in how buyers research and engage with vendor

Is there somewhere a collection of the best agent/coding harnesses for each models, especially open-source and local ones? In my opinion, th…

AgentsDGX agent

Is there somewhere a collection of the best agent/coding harnesses for each models, especially open-source and local ones? In my opinion, the biggest reason why people are struggling with open/local m

It’s not governance slowing down enterprise AI — it’s the lack of it, says Qlik executive

AgentsDGX agent

Enterprises are chasing AI models at a dizzying pace, yet the organizations pulling ahead are the ones that paused to build something less glamorous: a trusted, governed data foundation. Pressure is m

Learning to Focus and Precise Cropping: A Reinforcement Learning Framework with Information Gaps and Grounding Loss for MLLMs

AgentsDGX agent

arXiv:2603.27494v2 Announce Type: replace-cross Abstract: To enhance the perception and reasoning capabilities of multimodal large language models in complex visual scenes, recent research has introdu

MAVEN-T: Multi-Agent enVironment-aware Enhanced Neural Trajectory predictor with Reinforcement Learning

Model ReleasesDGX agent

arXiv:2604.10169v1 Announce Type: new Abstract: Trajectory prediction remains a critical yet challenging component in autonomous driving systems, requiring sophisticated reasoning capabilities while m

On Feedback Speed Control for a Planar Tracking

AgentsDGX agent

arXiv:2604.09795v1 Announce Type: cross Abstract: This paper investigates a planar tracking problem between a leader and follower agent. We propose a novel feedback speed control law, paired with a co

PaperScope: A Multi-Modal Multi-Document Benchmark for Agentic Deep Research Across Massive Scientific Papers

Model ReleasesDGX agent

arXiv:2604.11307v1 Announce Type: new Abstract: Leveraging Multi-modal Large Language Models (MLLMs) to accelerate frontier scientific research is promising, yet how to rigorously evaluate such system

RTMC: Step-Level Credit Assignment via Rollout Trees

AgentsDGX agent

arXiv:2604.11037v1 Announce Type: cross Abstract: Multi-step agentic reinforcement learning benefits from fine-grained credit assignment, yet existing approaches offer limited options: critic-free met

Towards Autonomous Mechanistic Reasoning in Virtual Cells

AgentsDGX agent

arXiv:2604.11661v1 Announce Type: cross Abstract: Large language models (LLMs) have recently gained significant attention as a promising approach to accelerate scientific discovery. However, their app

Tracing the Roots: A Multi-Agent Framework for Uncovering Data Lineage in Post-Training LLMs

Model ReleasesDGX agent

arXiv:2604.10480v1 Announce Type: new Abstract: Post-training data plays a pivotal role in shaping the capabilities of Large Language Models (LLMs), yet datasets are often treated as isolated artifact

13 Apr 2026

just finished this. genuinely one of the most informative episodes I've heard in a while the future of engineering is changing faster than m…

AgentsDGX agent

just finished this. genuinely one of the most informative episodes I've heard in a while the future of engineering is changing faster than most people realize. and it's self-managed, self-improving ag

@LangChain literally has the best devrels. They put out such great technical and educational content. Their blogs are also amazing! Definite…

AgentsDGX agent

Harrison Chase, co-founder of LangChain, shared or engaged with a post on X (formerly Twitter) praising LangChain's developer relations team for producing high-quality technical and educational conten

RIRF: Reasoning Image Restoration Framework

AgentsDGX agent

arXiv:2604.09511v1 Announce Type: new Abstract: Universal image restoration (UIR) aims to recover clean images from diverse and unknown degradations using a unified model. Existing UIR methods primari

12 Apr 2026

“Above the api line” Good phrase

AgentsDGX agent

“Above the api line” Good phrase GBrain is my attempt to be in control of my own personal AI that could become my intentionally designed cognitive armor Open source open prompts means you aren’t under

harness engineering > prompt engineering

AgentsDGX agent

Harrison Chase, co-founder of LangChain, has shared perspectives on X (formerly Twitter) distinguishing 'harness engineering' from traditional prompt engineering, suggesting a conceptual evolution in

if you don't own your harness, you don't own your memory this is so true. even though Codex is an open source, it generates an encrypted com…

Model ReleasesDGX agent

if you don't own your harness, you don't own your memory this is so true. even though Codex is an open source, it generates an encrypted compaction summary (that is not usable outside of the OpenAI ec

Only OG's know @NousResearch had bots back in 2024. This is when models were not capable. They've tried to solve this problem every way poss…

AgentsDGX agent

Only OG's know @NousResearch had bots back in 2024. This is when models were not capable. They've tried to solve this problem every way possible. Even @karan4d was exploring such ideas acitvely, @max_

The differentiating factor between a prototype and an autonomous system is no longer solely the underlying model weights, but the sophistica…

AgentsDGX agent

The differentiating factor between a prototype and an autonomous system is no longer solely the underlying model weights, but the sophistication of the orchestration layer and its capacity for continu

Welcome to Agents Week

IndustryDGX agent

Cloudflare's mission has always been to help build a better Internet. Sometimes that means building for the Internet as it exists. Sometimes it means building for the Internet as it's about to become.

11 Apr 2026

@hwchase17 @sarahwooders like a the old children's school flipbook - each one wakes up, takes in the context(memory) via harness does its th…

AgentsDGX agent

@hwchase17 @sarahwooders like a the old children's school flipbook - each one wakes up, takes in the context(memory) via harness does its thing, then rests.... next flip book page, the harness manages

@Xiaomi https://portal.nousresearch.com

ResearchDGX agent

Nous Research partnered with Xiaomi to integrate Xiaomi's MiMo V2 Pro inference model into its Hermes Agent framework, making it accessible via the Nous Portal for a two-week free trial period. The...

10 Apr 2026

Accelerating data curation with Google Data Cloud

Model ReleasesDGX agent

In the enterprise landscape, data is often highly fragmented across multiple source systems. Data curation is the process of organizing, cleaning, and enriching raw data to transform it into high-qual

Blockchain and AI: Securing Intelligent Networks for the Future

AgentsDGX agent

arXiv:2604.06323v2 Announce Type: cross Abstract: Blockchain and artificial intelligence (AI) are increasingly proposed together for securing intelligent networks, but the literature remains fragmente

Data Studio returns as new home for Data Cloud assets

AgentsDGX agent

In today's data-rich environment, organizations possess vast amounts of information. Yet, bridging the gap between that data and the users who need to make daily, informed decisions remains a challeng

Fighting AI with AI: AI-Agent Augmented DNS Blocking of LLM Services during Student Evaluations

Model ReleasesDGX agent

arXiv:2604.02360v1 Announce Type: cross Abstract: The transformative potential of large language models (LLMs) in education, such as improving accessibility and personalized learning, is being eclipse

GLM-5.1 is now available in Windsurf! Try it out and let us know what you think

AgentsDGX agent

GLM-5.1, developed by Z.ai (formerly Zhipu AI), is Z.ai's next-generation flagship model for agentic engineering, featuring 'significantly stronger coding capabilities than its predecessor,' bette...

How Waldium made a blog platform work for humans and AI alike

ToolsDGX agent

Waldium is a YC-backed, two-person agentic CMS that automates content research and creation, and gives every customer blog its own Model Context Protocol (MCP) server endpoint so AI agents can quer...

MedVR: Annotation-Free Medical Visual Reasoning via Agentic Reinforcement Learning

SafetyDGX agent

arXiv:2604.08203v1 Announce Type: new Abstract: Medical Vision-Language Models (VLMs) hold immense promise for complex clinical tasks, but their reasoning capabilities are often constrained by text-on

Mitigating Visual Context Degradation in Large Multimodal Models: A Training-Free Decoupled Agentic Framework

Model ReleasesDGX agent

arXiv:2509.23322v2 Announce Type: replace Abstract: With the continuous expansion of Large Language Models (LLMs) and advances in reinforcement learning, LLMs have demonstrated exceptional reasoning c

Ramp is setting the gold standard for AI usage for any company that’s not OpenAI/Anthropic If you’re not tokenmaxxing you’re falling behind

AgentsDGX agent

Ramp has launched **AI Spend Intelligence**, a new product that gives finance teams token-level visibility into AI costs across providers like OpenAI and Anthropic. The New York-based company pull...

The Planetary Cost of AI Acceleration, Part II: The 10th Planetary Boundary and the 6.5-Year Countdown

AgentsDGX agent

arXiv:2604.04956v2 Announce Type: replace-cross Abstract: The recent, super-exponential scaling of autonomous Large Language Model (LLM) agents signals a broader, fundamental paradigm shift from machi

We have @SamMorrowDrums talking about the challenges of building and scaling the @github MCP server at @aiDotEngineer 👀

AgentsDGX agent

Sam Morrow, a Senior Software Engineer at GitHub on the Copilot Agent Services team, spoke at the ai.engineer conference about the challenges of building and scaling the GitHub MCP Server — an open...

Wrapped up the 3-day AI Engineer conference in London. Key takeaways: Planning & verification > implementation. Humans define what to build …

Model ReleasesDGX agent

Wrapped up the 3-day AI Engineer conference in London. Key takeaways: Planning & verification > implementation. Humans define what to build and validate it works. AI handles the coding. The harness ma

9 Apr 2026

What’s new in Microsoft Foundry | March 2026

Model ReleasesDGX agent

March ships Foundry Agent Service GA with private networking, GPT-5.4 and GPT-5.4 Mini, Priority Processing, Phi-4 Reasoning Vision, SDK 2.0 GA across Python, JS/TS, Java, and .NET, Fireworks AI and

8 Apr 2026

'Claude Code isn't magic. The harness layer is just software, and software is something any dev can shape to fit how they want to work.' Che…

Model ReleasesDGX agent

'Claude Code isn't magic. The harness layer is just software, and software is something any dev can shape to fit how they want to work.' Check out @Hacubu’s practical guide to building a custom agent

7 Apr 2026

🤖 【2026最新】1秒で理解!8時間自律稼働する「GLM-5.1」がLinuxをゼロから構築。オープンソースが世界を制す極限全貌 【2026最新】1秒で理解!8時間自律稼働し、Linuxデスクトップをゼロから完成させる衝撃のオープンソースAI「GLM-5.1」が登場。SWE-…

AgentsDGX agent

🤖 【2026最新】1秒で理解!8時間自律稼働する「GLM-5.1」がLinuxをゼロから構築。オープンソースが世界を制す極限全貌 【2026最新】1秒で理解!8時間自律稼働し、Linuxデスクトップをゼロから完成させる衝撃のオープンソースAI「GLM-5.1」が登場。SWE-Bench Proで世界3位(OSS首位)を記録し、コーディングとエージェント性能を極限まで高めたその実態と、エンジ… h

14 Aug 2026

AaLLM: An End-to-End Analog Circuit Design Framework from Topology Generation to Sizing Using Large Language Models

AgentsDGX agent

arXiv:2608.13472v1 Announce Type: cross Abstract: Analog circuit design is a time-consuming, iterative process in a nonlinear and high-dimensional design space that relies heavily on expert intuition.

QuoteBench: How Matched Scores Can Hide Command-Path Failures

Model ReleasesDGX agent

arXiv:2608.13547v1 Announce Type: new Abstract: LLM coding agents issue Bash commands through interfaces that may serialize, wrap, and reparse model output. Matched execution scores alone cannot disti

Qwen Live EP2 - Agent First: Multimodal Gets to Work - Streamed live 11 hours ago

Model ReleasesDGX agent

Possibly best way to spend current hour before grabbing 27B model(Instead of opening duplicate repeated 27B posts here). I tried to grab summary(thought of including in this thread) of this video usin

13 Aug 2026

While waiting for the release of Qwen3.8-27B, let's try to guess what will happen

AgentsDGX agent

They highlighted 3 things on countdown page: VLM, Agentic Improvements, and Think mode. What improvements do you expect? Reply here! Personally, I want to meet a sage who has attained enlightenment. T

12 Aug 2026

A model is only as good as its data, and we’ve long since exhausted the internet. From here on out, model progress is gated by data producti…

AgentsDGX agent

A model is only as good as its data, and we’ve long since exhausted the internet. From here on out, model progress is gated by data production. @mercor_ai’s @BrendanFoody joined us at our Sovereign AI

I believe the demand for compute is going to go up faster than the supply of compute, so the price of compute is going to increase significa…

AgentsDGX agent

I believe the demand for compute is going to go up faster than the supply of compute, so the price of compute is going to increase significantly in the future, perhaps as much as 10x in the next few y

MT-PingEval: Evaluating Multi-Turn Collaboration with Private Information Games

AgentsDGX agent

arXiv:2602.24188v2 Announce Type: replace Abstract: We present a scalable and verifiable methodology for evaluating language models in multi-turn interactions, using a suite of collaborative games tha

What Iterated Self-Feeding Probes of Language Models Measure, and a test that separates the construction from the model

AgentsDGX agent

arXiv:2608.10986v1 Announce Type: new Abstract: A growing class of methods probes a language model by feeding it its own output: self-consistency, iterated refinement, agentic loops. We ask what such

11 Aug 2026

Agentic Harnesses: LLM-Driven Verification Layers for Robot Autonomy

SafetyDGX agent

arXiv:2608.09857v1 Announce Type: cross Abstract: Advances in advanced artificial intelligence tools have sparked research in robot autonomy, but the development of such systems has largely focused on

Agentic Visual Reasoning in Whole-Slide Pathology Images via Active Perception

SafetyDGX agent

arXiv:2608.08648v1 Announce Type: new Abstract: Whole-slide visual reasoning requires identifying sparse diagnostic evidence in gigapixel pathology slides and integrating observations across spatial s

← Previous
1…149150151152153…300
Next →