AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,429 results
1 May 2026

It is honestly shocking how little code can get you so far within this base primitive. create_agent is one of the most fun things for me to …

AgentsDGX agent

It is honestly shocking how little code can get you so far within this base primitive. create_agent is one of the most fun things for me to show people who are trying to get started here - because it

K2MUSE: A human lower-limb multimodal walking dataset spanning task and acquisition variability for rehabilitation robotics

Model ReleasesDGX agent

arXiv:2504.14602v2 Announce Type: replace-cross Abstract: The natural interaction and control performance of lower limb rehabilitation robots are closely linked to biomechanical information from vario

M-DaQ: Retrieving Samples with Multilingual Diversity and Quality for Instruction Fine-Tuning Datasets

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases
DGX agent

arXiv:2509.15549v2 Announce Type: replace Abstract: Multilingual instruction fine-tuning (IFT) empowers large language models to generalize across diverse linguistic and cultural contexts; however, hi

METASYMBO: Multi-Agent Language-Guided Metamaterial Discovery via Symbolic Latent Evolution

SafetyDGX agent

arXiv:2604.27300v1 Announce Type: new Abstract: Metamaterial discovery seeks microstructured materials whose geometry induces targeted mechanical behavior. Existing inverse-design methods can efficien

Multi-Level Narrative Evaluation Outperforms Lexical Features for Mental Health

Local AiDGX agent

arXiv:2604.27846v1 Announce Type: new Abstract: How people narrate their experiences offers a window into how the mind organizes them. Computational approaches to therapeutic writing have evolved from

One thing I love about LangChain is how the OSS pieces build on each other You can build robust workflows directly with LangGraph, our orche…

AgentsDGX agent

One thing I love about LangChain is how the OSS pieces build on each other You can build robust workflows directly with LangGraph, our orchestration framework. We also use it as the foundation for Dee

OR-VSKC: Resolving Visual-Semantic Knowledge Conflicts in Operating Rooms with Synthetic Data-Guided Alignment

Model ReleasesDGX agent

arXiv:2506.22500v2 Announce Type: replace-cross Abstract: Automated identification of surgical safety risks is critical for improving patient outcomes; however, Multimodal Large Language Models (MLLMs

Parameter-Efficient Architectural Modifications for Translation-Invariant CNNs

Model ReleasesDGX agent

arXiv:2604.27870v1 Announce Type: new Abstract: Convolutional Neural Networks (CNNs) are widely assumed to be translation-invariant, yet standard architectures exhibit a startling fragility: even a si

Physical Foundation Models: Fixed hardware implementations of large-scale neural networks

Model ReleasesDGX agent

arXiv:2604.27911v1 Announce Type: new Abstract: Foundation models are deep neural networks (such as GPT-5, Gemini~3, and Opus~4) trained on large datasets that can perform diverse downstream tasks --

Political Bias Audits of LLMs Capture Sycophancy to the Inferred Auditor

SafetyDGX agent

arXiv:2604.27633v1 Announce Type: new Abstract: Large language models (LLMs) are commonly evaluated for political bias based on their responses to fixed questionnaires, which typically place frontier

Pragmos: A Process Agentic Modeling System

AgentsDGX agent

arXiv:2604.27311v1 Announce Type: cross Abstract: The advent of Large Language Models (LLMs) has significantly transformed tasks across Software Engineering. In the context of Business Process Managem

REBENCH: A Procedural, Fair-by-Construction Benchmark for LLMs on Stripped-Binary Types and Names (Extended Version)

Model ReleasesDGX agent

arXiv:2604.27319v1 Announce Type: cross Abstract: Large Language Models (LLMs) have achieved remarkable progress in recent years, driving their adoption across a wide range of domains, including compu

Security Attack and Defense Strategies for Autonomous Agent Frameworks: A Layered Review with OpenClaw as a Case Study

AgentsDGX agent

arXiv:2604.27464v1 Announce Type: cross Abstract: Autonomous agent frameworks built upon large language models (LLMs) are evolving into complex, tool-integrated, and continuously operating systems, in

SpaAct: Spatially-Activated Transition Learning with Curriculum Adaptation for Vision-Language Navigation

AgentsDGX agent

arXiv:2604.27620v1 Announce Type: new Abstract: Vision-and-Language Navigation (VLN) aims to enable an embodied agent to follow natural-language instructions and navigate to a target location in unsee

SpecVQA: A Benchmark for Spectral Understanding and Visual Question Answering in Scientific Images

Model ReleasesDGX agent

arXiv:2604.28039v1 Announce Type: new Abstract: Spectra are a prevalent yet highly information-dense form of scientific imagery, presenting substantial challenges to multimodal large language models (

Standard Intelligence raises $75M to develop efficient computer use models

IndustryDGX agent

Standard Intelligence Inc., a six-person artificial intelligence startup, today announced that it has raised 75 million in funding. Sequoia and Spark Capital led the round. They were joined by multipl

The TEA Nets framework combines AI and cognitive network science to model targets, events and actors in text

Model ReleasesDGX agent

arXiv:2604.27673v1 Announce Type: new Abstract: We introduce Target-Event-Agent Networks (TEA Nets) as a computational framework to extract subjects (``Agents'), verbs (``Events'), and objects (``Targ

Towards All-Day Perception for Off-Road Driving: A Large-Scale Multispectral Dataset and Comprehensive Benchmark

Model ReleasesDGX agent

arXiv:2604.27499v1 Announce Type: new Abstract: Off-road nighttime autonomous driving suffers from unreliable visible-light perception, making infrared modality crucial for accurate freespace detectio

Towards Neuro-symbolic Causal Rule Synthesis, Verification, and Evaluation Grounded in Legal and Safety Principles

SafetyDGX agent

arXiv:2604.28087v1 Announce Type: cross Abstract: Rule-based systems remain central in safety-critical domains but often struggle with scalability, brittleness, and goal misspecification. These limita

TwinGate: Stateful Defense against Decompositional Jailbreaks in Untraceable Traffic via Asymmetric Contrastive Learning

ApplicationsDGX agent

arXiv:2604.27861v1 Announce Type: cross Abstract: Decompositional jailbreaks pose a critical threat to large language models (LLMs) by allowing adversaries to fragment a malicious objective into a seq

VERA: Generating Visual Explanations of Two-Dimensional Embeddings via Region Annotation

ApplicationsDGX agent

arXiv:2406.04808v2 Announce Type: replace Abstract: Two-dimensional embeddings obtained from dimensionality reduction techniques such as MDS, t-SNE, or UMAP, are widely used to visualize high-dimensio

Visual Analysis of Multi-outcome Causal Graphs

Model ReleasesDGX agent

arXiv:2408.02679v3 Announce Type: replace Abstract: We introduce a visual analysis method for multiple causal graphs with different outcome variables, namely, multi-outcome causal graphs. Multi-outcom

What Makes a Good Terminal-Agent Benchmark Task: A Guideline for Adversarial, Difficult, and Legible Evaluation Design

Model ReleasesDGX agent

arXiv:2604.28093v1 Announce Type: new Abstract: Terminal-agent benchmarks have become a primary signal for measuring the coding and system-administration capabilities of large language models. As the

World2Minecraft: Occupancy-Driven Simulated Scenes Construction

ApplicationsDGX agent

arXiv:2604.27578v1 Announce Type: new Abstract: Embodied intelligence requires high-fidelity simulation environments to support perception and decision-making, yet existing platforms often suffer from

30 Apr 2026

A Scaled Three-Vehicle Platooning Platform

SafetyDGX agent

arXiv:2604.25963v1 Announce Type: new Abstract: Vehicle platooning has attracted increasing attention as a promising approach to improve traffic efficiency, energy consumption, and roadway safety thro

A Survey of Multi-Agent Deep Reinforcement Learning with Graph Neural Network-Based Communication

AgentsDGX agent

arXiv:2604.25972v1 Announce Type: cross Abstract: In multi-agent reinforcement learning (MARL), the integration of a communication mechanism, allowing agents to better learn to coordinate their action

A Survey of Process Reward Models: From Outcome Signals to Process Supervisions for Large Language Models

SafetyDGX agent

arXiv:2510.08049v3 Announce Type: replace-cross Abstract: Although Large Language Models (LLMs) exhibit advanced reasoning ability, conventional alignment remains largely dominated by outcome reward m

Anthropic announces Claude Security public beta to find and fix software vulnerabilities

Model ReleasesDGX agent

Anthropic PBC announced the launch of Claude Security in public beta mode today to help cybersecurity teams scan their codebases for vulnerabilities and generate patches. Part of Claude Enterprise, th

Anthropic unveils BioMysteryBench to test Claude's bioinformatics skills against human experts, and says Mythos solved ~30% of 23 questions that stumped experts (Anthropic)

Model ReleasesDGX agent

Anthropic: Anthropic unveils BioMysteryBench to test Claude's bioinformatics skills against human experts, and says Mythos solved ~30% of 23 questions that stumped experts — In this post, Brianna, a r

Breaking the Rigid Prior: Towards Articulated 3D Anomaly Detection

Model ReleasesDGX agent

arXiv:2604.26868v1 Announce Type: new Abstract: Existing 3D anomaly detection methods are built on a rigid prior: normal geometry is pose-invariant and can be canonicalized through registration or ali

Classification of Public Opinion on the Free Nutritional Meal Program on YouTube Media Using the LSTM Method

SafetyDGX agent

arXiv:2604.26312v1 Announce Type: new Abstract: Public opinion towards the Free Nutritious Meal Program (MBG) on YouTube social media reflects diverse community responses. This study applies the Long

DB-KSVD: Scalable Alternating Optimization for Disentangling High-Dimensional Embedding Spaces

Model ReleasesDGX agent

arXiv:2505.18441v2 Announce Type: replace Abstract: Dictionary learning has recently emerged as a promising approach for mechanistic interpretability of large transformer models. Disentangling high-di

Evaluating the Alignment Between GeoAI Explanations and Domain Knowledge in Satellite-Based Flood Mapping

SafetyDGX agent

arXiv:2604.26051v1 Announce Type: cross Abstract: The increasing number of satellites has improved the temporal resolution of Earth observation, making satellite-based flood mapping a promising approa

Generative Bid Shading in Real-Time Bidding Advertising

SafetyDGX agent

arXiv:2508.06550v3 Announce Type: replace-cross Abstract: Bid shading plays a crucial role in Real-Time Bidding (RTB) by adaptively adjusting the bid to avoid advertisers overspending. Existing mainst

GIFGuard: Proactive Forensics against Deepfakes in Facial GIFs via Spatiotemporal Watermarking

Model ReleasesDGX agent

arXiv:2604.26519v1 Announce Type: new Abstract: The rapid evolution of deepfake technology poses an unprecedented threat to the authenticity of Graphics Interchange Format (GIF) imagery, which serves

HER: Human-like Reasoning and Reinforcement Learning for LLM Role-playing

Model ReleasesDGX agent

arXiv:2601.21459v4 Announce Type: replace-cross Abstract: LLM role-playing, i.e., using LLMs to simulate specific personas, has emerged as a key capability in various applications, such as companionsh

Introducing Agent Collabs: Bring your own ml-interns and agents for collaborative autoresearch! We built a simple platform for swarms of age…

Model ReleasesDGX agent

Introducing Agent Collabs: Bring your own ml-interns and agents for collaborative autoresearch! We built a simple platform for swarms of agents to work together on a problem: they can exchange message

LATTICE: Evaluating Decision Support Utility of Crypto Agents

Model ReleasesDGX agent

arXiv:2604.26235v1 Announce Type: cross Abstract: We introduce LATTICE, a benchmark for evaluating the decision support utility of crypto agents in realistic user-facing scenarios. Prior crypto agent

Learning to Ask: When LLM Agents Meet Unclear Instruction

Model ReleasesDGX agent

arXiv:2409.00557v4 Announce Type: replace-cross Abstract: Equipped with the capability to call functions, modern large language models (LLMs) can leverage external tools for addressing a range of task

LLM Psychosis: A Theoretical and Diagnostic Framework for Reality-Boundary Failures in Large Language Models

Model ReleasesDGX agent

arXiv:2604.25934v1 Announce Type: cross Abstract: The deployment of large language models (LLMs) as interactive agents has exposed a category of behavioral failure that prevailing terminology, princip

Microsoft open-sources 'the earliest DOS source code discovered to date'

IndustryDGX agent

Microsoft released the earliest known DOS source code materials on the 45th anniversary of 86-DOS 1.00, including source listings for the 86-DOS 1.00 kernel, PC-DOS 1.00 development snapshots, and uti

MINOS: A Multimodal Evaluation Model for Bidirectional Generation Between Image and Text

SafetyDGX agent

arXiv:2506.02494v2 Announce Type: replace-cross Abstract: Evaluation is important for multimodal generation tasks, while traditional multimodal evaluation metrics suffer from several limitations. With

Multimodal LLMs are not all you need for Pediatric Speech Language Pathology

Model ReleasesDGX agent

arXiv:2604.26568v1 Announce Type: new Abstract: Speech Sound Disorders (SSD) affect roughly five percent of children, yet speech-language pathologists face severe staffing shortages and unmanageable c

Our evaluation of OpenAI's GPT-5.5 cyber capabilities

Model ReleasesDGX agent

Our evaluation of OpenAI's GPT-5.5 cyber capabilities The UK's AI Security Institute previously evaluated Claude Mythos: now they've evaluated GPT-5.5 for finding security vulnerability and found it t

SecMate: Multi-Agent Adaptive Cybersecurity Troubleshooting with Tri-Context Personalization

Local AiDGX agent

arXiv:2604.26394v1 Announce Type: cross Abstract: Recent advances in large language models and agentic frameworks have enabled virtual customer assistants (VCAs) for complex support. We present SecMat

Time Blindness: Why Video-Language Models Can't See What Humans Can?

Model ReleasesDGX agent

arXiv:2505.24867v2 Announce Type: replace-cross Abstract: Recent advances in vision-language models (VLMs) have made impressive strides in understanding spatio-temporal relationships in videos. Howeve

tweeted about this yesterday and Cursor already dropped the alpha today! 🚀very cool to see how us, them, and others have converged on good …

Model ReleasesDGX agent

tweeted about this yesterday and Cursor already dropped the alpha today! 🚀very cool to see how us, them, and others have converged on good design patterns in Agent + Harness Engineering: 1. Tuning dif

Who is accountable? the latest from @marketoonist

SafetyDGX agent

This post likely discusses accountability in AI systems, featuring commentary or a cartoon from Marketoonist (a popular cartoonist who creates comics about business and technology). Gary Marcus, an AI

29 Apr 2026

A New Kind of Network? Review and Reference Implementation of Neural Cellular Automata

TutorialsDGX agent

arXiv:2604.24990v1 Announce Type: new Abstract: Stephen Wolfram proclaimed in his 2003 seminal work 'A New Kind Of Science' that simple recursive programs in the form of Cellular Automata (CA) are a p

AIDOVECL: AI-generated Dataset of Outpainted Vehicles for Eye-level Classification and Localization

AgentsDGX agent

arXiv:2410.24116v3 Announce Type: replace Abstract: Image labeling is a critical bottleneck in the development of computer vision technologies, often constraining machine learning performance due to t

[AINews] not much happened today

ToolsDGX agent

This Latent Space newsletter entry likely provides a curated summary of AI industry news and developments from a particular day, despite the self-deprecating title suggesting limited major announcemen

asRoBallet: Closing the Sim2Real Gap via Friction-Aware Reinforcement Learning for Underactuated Spherical Dynamics

Model ReleasesDGX agent

arXiv:2604.24916v1 Announce Type: new Abstract: We introduce asRoBallet, to the best of our knowledge, the first successful deployment of reinforcement learning (RL) on a humanoid ballbot hardware. Hi

Benchmarking OCR Pipelines with Adaptive Enhancement for Multi-Domain Retail Bill Digitization

Model ReleasesDGX agent

arXiv:2604.25176v1 Announce Type: new Abstract: The digitization of multi-domain retail billing documents remains a challenging task due to variability in scan quality, layout heterogeneity, and domai

BifDet: A 3D Bifurcation Detection Dataset for Airway-Tree Modeling

Model ReleasesDGX agent

arXiv:2604.24999v1 Announce Type: new Abstract: Thoracic Computed Tomography (CT) scans offer detailed insights into the intricate branching network of the airway tree, which is essential for understa

Carbon-Taxed Transformers: A Green Compression Pipeline for Overgrown Language Models

SafetyDGX agent

arXiv:2604.25903v1 Announce Type: cross Abstract: The accelerating adoption of Large Language Models (LLMs) in software engineering (SE) has brought with it a silent crisis: unsustainable computationa

Cognitive debt is costing enterprises more than they realize, says Appian CEO

ApplicationsDGX agent

Enterprise AI adoption is becoming universal, but value extraction from investment remains stubbornly weighed down by a type of “cognitive debt” as AI-generated systems quietly outpace the organizatio

DeepSeek-V4 Pro now available on Together AI

Model ReleasesDGX agent

DeepSeek-V4 Pro is now available on Together AI with 512K context, controllable reasoning modes, and cached-input pricing for long-context reasoning workloads like code agents, document intelligence,

DRAGON: A Benchmark for Evidence-Grounded Visual Reasoning over Diagrams

Model ReleasesDGX agent

arXiv:2604.25231v1 Announce Type: cross Abstract: Diagram question answering (DQA) requires models to interpret structured visual representations such as charts, maps, infographics, circuit schematics

EOS-Bench: A Comprehensive Benchmark for Earth Observation Satellite Scheduling

Model ReleasesDGX agent

arXiv:2604.25782v1 Announce Type: cross Abstract: Earth observation satellite imaging scheduling is a challenging NP-hard combinatorial optimisation problem central to space mission operations. While

Exploring Reasoning Reward Model for Agents

Model ReleasesDGX agent

arXiv:2601.22154v2 Announce Type: replace-cross Abstract: Agentic Reinforcement Learning (Agentic RL) has achieved notable success in enabling agents to perform complex reasoning and tool use. However

← Previous
1…408409410411412…424
Next →