AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,958 results
28 Apr 2026

Scheming Ability in LLM-to-LLM Strategic Interactions

Model ReleasesDGX agent

arXiv:2510.12826v2 Announce Type: replace-cross Abstract: As large language model (LLM) agents are deployed autonomously in diverse contexts, evaluating their capacity for strategic deception becomes

27 Apr 2026

Cross-Stage Coherence in Hierarchical Driving VQA: Explicit Baselines and Learned Gated Context Projectors

AgentsDGX agent

arXiv:2604.22560v1 Announce Type: cross Abstract: Graph Visual Question Answering (GVQA) for autonomous driving organizes reasoning into ordered stages, namely Perception, Prediction, and Planning, wh

Rethinking Publication: A Certification Framework for AI-Enabled Research

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases
DGX agent

arXiv:2604.22026v1 Announce Type: new Abstract: AI research pipelines now produce a growing share of publishable academic output, including work that meets existing peer-review standards for quality a

26 Apr 2026

I actually switched my personal Claude subscription to this (currently using Mimo v2)

Model ReleasesDGX agent

I actually switched my personal Claude subscription to this (currently using Mimo v2) Nous Portal offers everything you need to build with Hermes Agent in one easy subscription: → 300+ models from eve

NEW paper from Alibaba. A 30B MoE with only 3B active params matches Qwen3-235B on real tool-use workloads. AgenticQwen-30B-A3B: 50.2 averag…

Model ReleasesDGX agent

NEW paper from Alibaba. A 30B MoE with only 3B active params matches Qwen3-235B on real tool-use workloads. AgenticQwen-30B-A3B: 50.2 average on TAU-2 + BFCL-V4 Multi-Turn. AgenticQwen-8B: 47.4. Both

25 Apr 2026

Can’t wait for @LangChain Interrupt in two weeks! Best AI dev conf this year

AgentsDGX agent

Harrison Chase, co-founder of LangChain, expressed enthusiasm about an upcoming LangChain Interrupt conference event scheduled for two weeks from his post date. He characterized it as the best AI deve

wow another engineer on the “code is not cheap” train https://x.com/mattzcarey/status/2048102657512935626?s=46

ToolsDGX agent

wow another engineer on the “code is not cheap” train https://x.com/mattzcarey/status/2048102657512935626?s=46 My talk at AI Engineer “Every API is a Tool for Agents” is out on YouTube https://youtu.b

24 Apr 2026

Learning Reasoning Reward Models from Expert Demonstration via Inverse Reinforcement Learning

Local AiDGX agent

arXiv:2510.01857v3 Announce Type: replace Abstract: Current approaches to improving reasoning in large language models (LLMs) primarily rely on either supervised fine-tuning (SFT) over expert traces o

Probably the most useful advice I can give you if you run an ad agency, a marketing company, a production studio, or anything in that space.…

ApplicationsDGX agent

Probably the most useful advice I can give you if you run an ad agency, a marketing company, a production studio, or anything in that space. Hear me out. I’ve spent a decade talking to teams about thi

Task-specific Subnetwork Discovery in Reinforcement Learning for Autonomous Underwater Navigation

SafetyDGX agent

arXiv:2604.21640v1 Announce Type: cross Abstract: Autonomous underwater vehicles are required to perform multiple tasks adaptively and in an explainable manner under dynamic, uncertain conditions and

23 Apr 2026

Claude Code spend had gotten to $10.95M runrate peak at SemiAnalysis But then Opus 4.7 saved me. More token effecient for tasks, smarter, an…

Model ReleasesDGX agent

Claude Code spend had gotten to $10.95M runrate peak at SemiAnalysis But then Opus 4.7 saved me. More token effecient for tasks, smarter, and no fast mode. Thank you @AnthropicAI You saved me from ban

New Max Agency! 🎙️ @hwchase17 🎙️ @ListenLabs Co-Founder/CTO @florian_jue

AgentsDGX agent

New Max Agency! 🎙️ @hwchase17 🎙️ @ListenLabs Co-Founder/CTO @florian_jue 🎙️ Talked to @ListenLabs co-founder + CTO @florian_jue in the latest Max Agency. Really enjoyed hearing about the architectural

SweRank: Software Issue Localization with Code Ranking

Model ReleasesDGX agent

arXiv:2505.07849v2 Announce Type: replace-cross Abstract: Software issue localization, the task of identifying the precise code locations (files, classes, or functions) relevant to a natural language

Three shifts in the AI stack today: - OpenAI ships GPT-5.5 as a mid-cycle drop before its IPO - Hugging Face's ML Intern beats Claude Code a…

Model ReleasesDGX agent

Three shifts in the AI stack today: - OpenAI ships GPT-5.5 as a mid-cycle drop before its IPO - Hugging Face's ML Intern beats Claude Code and Codex on research - Pliny used Claude Opus 4.7 to jailbre

22 Apr 2026

Mitigating Judgment Preference Bias in Large Language Models through Group-Based Polling

SafetyDGX agent

arXiv:2510.08145v2 Announce Type: replace Abstract: Large Language Models (LLMs) as automatic evaluators, commonly referred to as LLM-as-a-Judge, have also attracted growing attention. This approach p

Small and midsize businesses jumpstart their AI transformations with Gemini Enterprise

Model ReleasesDGX agent

Small businesses are the backbone of the global economy. With 400 million SMBs worldwide and 36 million in the U.S. alone, they provide 50% of global employment. Now, with Google Cloud AI, they’re sca

21 Apr 2026

Decoding AI Tutor Effects for Educational Measurement: Temporal, Multi-Outcome, and Behavior-Cognitive Analysis

SafetyDGX agent

arXiv:2604.16366v1 Announce Type: cross Abstract: Artificial intelligence (AI) tutors have become increasingly popular in learning environments. In this study, we propose an AI agent prototype framewo

🔒`deepagents deploy` now supports custom auth: turn a single deployment into a multi-tenant platform. per-user resource and thread isolatio…

ApplicationsDGX agent

🔒`deepagents deploy` now supports custom auth: turn a single deployment into a multi-tenant platform. per-user resource and thread isolation, access control, and your own auth provider, all without sp

Heterogeneous Self-Play for Realistic Highway Traffic Simulation

SafetyDGX agent

arXiv:2604.16406v1 Announce Type: cross Abstract: Realistic highway simulation is critical for scalable safety evaluation of autonomous vehicles, particularly for interactions that are too rare to stu

I’ve been using ml-intern for a while, and it genuinely changed my workflow. It's super good at: - Model/Dataset discovery. - Post-Training …

HardwareDGX agent

I’ve been using ml-intern for a while, and it genuinely changed my workflow. It's super good at: - Model/Dataset discovery. - Post-Training setup iteration. - Data processing workflows. Huge shoutout

LookasideVLN: Direction-Aware Aerial Vision-and-Language Navigation

AgentsDGX agent

arXiv:2604.17190v1 Announce Type: new Abstract: Aerial Vision-and-Language Navigation (Aerial VLN) enables unmanned aerial vehicles (UAVs) to follow natural language instructions and navigate complex

MoCo: A One-Stop Shop for Model Collaboration Research

SafetyDGX agent

arXiv:2601.21257v2 Announce Type: replace Abstract: Advancing beyond single monolithic language models (LMs), recent research increasingly recognizes the importance of model collaboration, where multi

Moving past bots vs. humans

IndustryDGX agent

As AI assistants and privacy proxies challenge the capabilities of traditional bot detection, the Web needs new models for accountability. We believe that control should remain with the client, and th

OneVL: One-Step Latent Reasoning and Planning with Vision-Language Explanation

AgentsDGX agent

arXiv:2604.18486v1 Announce Type: cross Abstract: Chain-of-Thought (CoT) reasoning has become a powerful driver of trajectory prediction in VLA-based autonomous driving, yet its autoregressive nature

Tool Learning Needs Nothing More Than a Free 8B Language Model

ResearchDGX agent

arXiv:2604.17739v1 Announce Type: cross Abstract: Reinforcement learning (RL) has become a prevalent paradigm for training tool calling agents, which typically requires online interactive environments

20 Apr 2026

Aerial Multi-Functional RIS in Fluid Antennas-Aided Full-Duplex Networks: A Self-Optimized Hybrid Deep Reinforcement Learning Approach

SafetyDGX agent

arXiv:2604.14309v2 Announce Type: replace-cross Abstract: To address high data traffic demands of sixth-generation (6G) networks, this paper proposes a novel architecture that integrates autonomous ae

Chinese tech workers are starting to train their AI doubles–and pushing back

ResearchDGX agent

Tech workers in China are being instructed by their bosses to train AI agents to replace them—and it’s prompting a wave of soul-searching among otherwise enthusiastic early adopters. Earlier this mont

18 Apr 2026

i am forced to recommend viewing of sunil's talk if i want to retain healthy knees https://www.youtube.com/live/O_IMsEg91g8?t=31662&is=WIMl-…

AgentsDGX agent

i am forced to recommend viewing of sunil's talk if i want to retain healthy knees https://www.youtube.com/live/O_IMsEg91g8?t=31662&is=WIMl-Di3-Xi6gtse You should watch these 2 talks from AI Engineer

🆕The Friction is Your Judgment https://www.youtube.com/watch?v=_Zcw_sVF6hU @mitsuhiko (creator of Flask) and Cristina Poncela Cubeiro (AI-n…

AgentsDGX agent

🆕The Friction is Your Judgment https://www.youtube.com/watch?v=_Zcw_sVF6hU @mitsuhiko (creator of Flask) and Cristina Poncela Cubeiro (AI-native Engineer) break down what goes WRONG when you 'ship wit

17 Apr 2026

Hey that’s me!

TutorialsDGX agent

Hey that’s me! 🆕 Harness Engineering: How to Build Software When Humans Steer, Agents Execute https://www.youtube.com/watch?v=am_oeAoUhew @_lopopolo is one of the emerging class of token billionaires

@_lopopolo @aiDotEngineer @badlogicgames Harnes Engineering talk: https://x.com/aiDotEngineer/status/2044937501593460773?s=20

TutorialsDGX agent

@_lopopolo @aiDotEngineer @badlogicgames Harnes Engineering talk: https://x.com/aiDotEngineer/status/2044937501593460773?s=20 🆕 Harness Engineering: How to Build Software When Humans Steer, Agents Exe

NLP needs Diversity outside of 'Diversity'

SafetyDGX agent

arXiv:2604.14595v1 Announce Type: new Abstract: This position paper argues that recent progress with diversity in NLP is disproportionately concentrated on a small number of areas surrounding fairness

TableNet A Large-Scale Table Dataset with LLM-Powered Autonomous

AgentsDGX agent

arXiv:2604.13041v1 Announce Type: cross Abstract: Table Structure Recognition (TSR) requires the logical reasoning ability of large language models (LLMs) to handle complex table layouts, but current

World-Value-Action Model: Implicit Planning for Vision-Language-Action Systems

TutorialsDGX agent

arXiv:2604.14732v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have emerged as a promising paradigm for building embodied agents that ground perception and language into action.

16 Apr 2026

Beyond State Consistency: Behavior Consistency in Text-Based World Models

SafetyDGX agent

arXiv:2604.13824v1 Announce Type: new Abstract: World models have been emerging as critical components for assessing the consequences of actions generated by interactive agents in online planning and

C^2T: Captioning-Structure and LLM-Aligned Common-Sense Reward Learning for Traffic--Vehicle Coordination

SafetyDGX agent

arXiv:2604.13098v1 Announce Type: cross Abstract: State-of-the-art (SOTA) urban traffic control increasingly employs Multi-Agent Reinforcement Learning (MARL) to coordinate Traffic Light Controllers (

Red Skills or Blue Skills? A Dive Into Skills Published on ClawHub

Model ReleasesDGX agent

arXiv:2604.13064v1 Announce Type: new Abstract: Skill ecosystems have emerged as an increasingly important layer in Large Language Model (LLM) agent systems, enabling reusable task packaging, public d

Training-Free Test-Time Contrastive Learning for Large Language Models

AgentsDGX agent

arXiv:2604.13552v1 Announce Type: new Abstract: Large language models (LLMs) demonstrate strong reasoning capabilities, but their performance often degrades under distribution shift. Existing test-tim

We're proud to be on the @Forbes AI 50 list again! It reflects our focus on building secure, sovereign AI that helps enterprises and governm…

TutorialsDGX agent

We're proud to be on the @Forbes AI 50 list again! It reflects our focus on building secure, sovereign AI that helps enterprises and governments put their data to work on their terms. Learn more: http

15 Apr 2026

A Comparison of Reinforcement Learning and Optimal Control Methods for Path Planning

SafetyDGX agent

arXiv:2604.12628v1 Announce Type: cross Abstract: Path-planning for autonomous vehicles in threat-laden environments is a fundamental challenge. While traditional optimal control methods can find idea

BLOSSOM: Block-wise Federated Learning Over Shared and Sparse Observed Modalities

AgentsDGX agent

arXiv:2603.27552v2 Announce Type: replace Abstract: Multimodal federated learning (FL) is essential for real-world applications such as autonomous systems and healthcare, where data is distributed acr

CompliBench: Benchmarking LLM Judges for Compliance Violation Detection in Dialogue Systems

Model ReleasesDGX agent

arXiv:2604.12312v1 Announce Type: new Abstract: As Large Language Models (LLMs) are increasingly deployed as task-oriented agents in enterprise environments, ensuring their strict adherence to complex

Cool stuff Google Cloud customers built, April edition: BMW big on SLMs, MLB’s Scout Insights AI, personalized resort experiences

Model ReleasesDGX agent

AI and cloud technology are reshaping every corner of every industry around the world. Without our customers, who are building the future on our platform, there would be no Google Cloud. In this regul

Incentivizing High-Quality Human Annotations with Golden Questions

SafetyDGX agent

arXiv:2505.19134v2 Announce Type: replace-cross Abstract: Human-annotated data plays a vital role in training large language models (LLMs), such as supervised fine-tuning and human preference alignmen

Thermodynamic Liquid Manifold Networks: Physics-Bounded Deep Learning for Solar Forecasting in Autonomous Off-Grid Microgrids

AgentsDGX agent

arXiv:2604.11909v1 Announce Type: cross Abstract: The stable operation of autonomous off-grid photovoltaic systems requires solar forecasting algorithms that respect atmospheric thermodynamics. Contem

14 Apr 2026

A Survey on Deep Learning Techniques for Action Anticipation

AgentsDGX agent

arXiv:2309.17257v2 Announce Type: replace Abstract: The ability to anticipate possible future human actions is essential for a wide range of applications, including autonomous driving and human-robot

ActorMind: Emulating Human Actor Reasoning for Speech Role-Playing

Model ReleasesDGX agent

arXiv:2604.11103v1 Announce Type: cross Abstract: Role-playing has garnered rising attention as it provides a strong foundation for human-machine interaction and facilitates sociological research. How

Characterizing Performance-Energy Trade-offs of Large Language Models in Multi-Request Workflows

HardwareDGX agent

arXiv:2604.09611v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used in applications forming multi-request workflows like document summarization, search-based copilots,

CheeseBench: Evaluating Large Language Models on Rodent Behavioral Neuroscience Paradigms

Model ReleasesDGX agent

arXiv:2604.10825v1 Announce Type: new Abstract: We introduce CheeseBench, a benchmark that evaluates large language models (LLMs) on nine classical behavioral neuroscience paradigms (Morris water maze

Energy-Efficient Federated Edge Learning For Small-Scale Datasets in Large IoT Networks

AgentsDGX agent

arXiv:2604.10662v1 Announce Type: new Abstract: Large-scale Internet of Things (IoT) networks enable intelligent services such as smart cities and autonomous driving, but often face resource constrain

GroupRank: A Groupwise Paradigm for Effective and Efficient Passage Reranking with LLMs

Model ReleasesDGX agent

arXiv:2511.11653v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have emerged as powerful tools for passage reranking in information retrieval, leveraging their superior reasonin

If an LLM Were a Character, Would It Know Its Own Story? Evaluating Lifelong Learning in LLMs

Model ReleasesDGX agent

arXiv:2503.23514v2 Announce Type: replace-cross Abstract: Large language models (LLMs) can carry out human-like dialogue, but unlike humans, they are stateless due to the superposition property. Howev

MeloTune: On-Device Arousal Learning and Peer-to-Peer Mood Coupling for Proactive Music Curation

Local AiDGX agent

arXiv:2604.10815v1 Announce Type: cross Abstract: MeloTune is an iPhone-deployed music agent that instantiates the Mesh Memory Protocol (MMP) and Symbolic-Vector Attention Fusion (SVAF) as a productio

NEW: CA gubernatorial candidate Tom Steyer (D) releases an immigration platform that is radically left of Gov. Newsom. It includes: - Abolis…

ApplicationsDGX agent

NEW: CA gubernatorial candidate Tom Steyer (D) releases an immigration platform that is radically left of Gov. Newsom. It includes: - Abolish ICE - Put ICE agents in jail & “treat them like the mob”.

Rebooting Microreboot: Architectural Support for Safe, Parallel Recovery in Microservice Systems

SafetyDGX agent

arXiv:2604.09963v1 Announce Type: cross Abstract: Microreboot enables fast recovery by restarting only the failing component, but in modern microservices naive restarts are unsafe: dense dependencies

Shuffling the Data, Stretching the Step-size: Sharper Bias in constant step-size SGD

SafetyDGX agent

arXiv:2604.10373v1 Announce Type: cross Abstract: From adversarial robustness to multi-agent learning, many machine learning tasks can be cast as finite-sum min-max optimization or, more generally, as

Unified Unsupervised and Sparsely-Supervised 3D Object Detection by Semantic Pseudo-Labeling and Prototype Learning

AgentsDGX agent

arXiv:2602.21484v2 Announce Type: replace Abstract: 3D object detection is essential for autonomous driving and robotic perception, yet its reliance on large-scale manually annotated data limits scala

Utilizing and Calibrating Hindsight Process Rewards via Reinforcement with Mutual Information Self-Evaluation

SafetyDGX agent

arXiv:2604.11611v1 Announce Type: new Abstract: To overcome the sparse reward challenge in reinforcement learning (RL) for agents based on large language models (LLMs), we propose Mutual Information S

We've been building something that doesn't fit in a wave. Coming soon

ToolsDGX agent

Windsurf, the AI-powered coding platform, has teased an upcoming product or feature announcement that is described as something that doesn't fit within their existing 'wave' release framework. The cry

13 Apr 2026

CSAttention: Centroid-Scoring Attention for Accelerating LLM Inference

HardwareDGX agent

arXiv:2604.08584v1 Announce Type: cross Abstract: Long-context LLMs increasingly rely on extended, reusable prefill prompts for agents and domain Q&A, pushing attention and KV-cache to become the domi

← Previous
1…193194195196197…300
Next →