AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,629 results
27 May 2026

CmIVTP: Cross-modal Interaction-based Vessel Trajectory Prediction for Maritime Intelligence

SafetyDGX agent

arXiv:2605.26524v1 Announce Type: cross Abstract: Maritime intelligent transportation systems (MITS) are essential for ensuring navigation safety and efficiency in busy waterways. However, accurate ve

Completely agree, @hwchase17 ! This is the meta-level breakthrough we've all been waiting for. Self-optimizing loops finally feel production…

SafetyDGX agent

Completely agree, @hwchase17 ! This is the meta-level breakthrough we've all been waiting for. Self-optimizing loops finally feel production-ready because LangSmith Engine turns evaluation from a manu

Deep-layer limit and stability analysis of the basic forward-backward-splitting induced network (II): learning problems

Model Releases
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2605.27133v1 Announce Type: cross Abstract: Deep unfolding neural networks derived from iterative optimization schemes and numerical ordinary/partial differential equations (ODEs/PDEs) have attr

Demystifying Video Reasoning

TutorialsDGX agent

arXiv:2603.16870v2 Announce Type: replace-cross Abstract: Recent advances in video generation have revealed an unexpected phenomenon: diffusion-based video models exhibit non-trivial reasoning capabil

Efficient On-policy Visual-RL via Stochastic Decoupled Policy Gradient

SafetyDGX agent

arXiv:2605.26478v1 Announce Type: cross Abstract: We present the stochastic decoupled policy gradient (SDPG), a lightweight visual reinforcement learning (RL) method that trains diverse visuomotor con

Enabling Extensible Embodied Capabilities with Tools

SafetyDGX agent

arXiv:2605.26637v1 Announce Type: new Abstract: Most existing embodied intelligence methods formulate perception, reasoning, planning, and control within a unified parameterized policy. Yet these capa

ENPMR-Bench: Benchmarking Proactive Memory Retrieval for Emotional Support Agents

Model ReleasesDGX agent

arXiv:2605.27240v1 Announce Type: new Abstract: Memory-augmented language agents are increasingly deployed in affective applications such as emotional support, where understanding and responding to us

Falcon-X: A Time Series Foundation Model for Heterogeneous Multivariate Modeling

Model ReleasesDGX agent

arXiv:2605.27286v1 Announce Type: cross Abstract: Time series foundation models (TSFMs) are transforming the forecasting paradigm through large-scale cross-domain pretraining. However, most existing T

I think Anthropic and OpenAI have found product-market fit

Model ReleasesDGX agent

Anthropic are strongly rumored to be about to have their first profitable quarter. Stories are circulating of companies surprised at how expensive their LLM bills are becoming from usage by their staf

Implementation of Big Data Analytics for Diabetes Management: Needs Assessment in the Rwanda Healthcare System

ApplicationsDGX agent

arXiv:2605.26786v1 Announce Type: cross Abstract: Diabetes is a chronic metabolic disease that can lead to serious health problems if not diagnosed and managed early. Big Data Analytics (BDA) and mach

ImViD: Immersive Volumetric Videos for Enhanced VR Engagement

Model ReleasesDGX agent

arXiv:2503.14359v2 Announce Type: replace Abstract: User engagement is greatly enhanced by fully immersive multi-modal experiences that combine visual and auditory stimuli. Consequently, the next fron

Intelligent Offloading in Vehicular Edge Computing: A Comprehensive Review of Deep Reinforcement Learning Approaches and Architectures

SafetyDGX agent

arXiv:2502.06963v3 Announce Type: replace-cross Abstract: The increasing complexity of Intelligent Transportation Systems (ITS) has led to significant interest in computational offloading to external

Is Agent Memory a Database? Rethinking Data Foundations for Long-Term AI Agent Memory

Local AiDGX agent

arXiv:2605.26252v1 Announce Type: new Abstract: Long-running AI agents need persistent memory. Memory supports learning across sessions, reduces repeated context injection, and enables auditing of pas

ITBench-AA: Frontier Models Score Below 50% on the First Benchmark for Agentic Enterprise IT Tasks — by Artificial Analysis and IBM

Model ReleasesDGX agent

ITBench-AA is a new benchmark developed by Artificial Analysis and IBM that evaluates frontier AI models on agentic enterprise IT tasks, with results showing that current leading models score below 50

Kandinsky 5.0: A Family of Foundation Models for Image and Video Generation

Model ReleasesDGX agent

arXiv:2511.14993v3 Announce Type: replace-cross Abstract: This report introduces Kandinsky 5.0, a family of state-of-the-art foundation models for high-resolution image and 10-second video synthesis.

LearnedCache: An eBPF-Integrated Perceptron-Based Eviction Policy for the Linux Page Cache

SafetyDGX agent

arXiv:2605.26168v1 Announce Type: cross Abstract: Linux is the foundation of the digital age, accounting for the majority of the cloud and mobile OS markets. Any device that runs Linux uses the Linux

Lessons from Penetration Tests on Large-Scale Agent Systems

AgentsDGX agent

arXiv:2605.27042v1 Announce Type: cross Abstract: As AI systems gain increasing autonomy and execution capability, the number of discovered security vulnerabilities continues to rise. However, many of

LLMs versus the Halting Problem: Characterizing Program Termination Reasoning

Model ReleasesDGX agent

arXiv:2601.18987v5 Announce Type: replace-cross Abstract: Determining whether a program terminates is a central problem in computer science. Turing's Halting Problem established termination as undecid

LongCat-Video-Avatar 1.5 Technical Report

Model ReleasesDGX agent

arXiv:2605.26486v1 Announce Type: new Abstract: Despite advances in audio-driven video generation, achieving commercial-grade stability remains challenging. We present LongCat-Video-Avatar 1.5, an upg

Managed Deep Agents is built for agents that need to work over long time horizons, use tools, preserve context, and produce artifacts. A few…

AgentsDGX agent

Managed Deep Agents is built for agents that need to work over long time horizons, use tools, preserve context, and produce artifacts. A few examples of what teams are building: ✅ Support + triage age

Managed deep agents is the easiest way to build and deploy long horizon agents Private preview, dm me if you want access

ApplicationsDGX agent

Managed deep agents is the easiest way to build and deploy long horizon agents Private preview, dm me if you want access Managed Deep Agents is built for agents that need to work over long time horizo

MULTISEISMO: A Multimodal Seismic Dataset and Model for Cross-Modal Seismic Understanding

Model ReleasesDGX agent

arXiv:2605.26320v1 Announce Type: cross Abstract: The application of generalist multimodal models (GMMs) to specialized scientific domains remains limited due to the scarcity of comprehensive domain-s

Neuro-Symbolic Verification of LLM Outputs for Data-Sensitive Domains (extended preprint)

SafetyDGX agent

arXiv:2605.26942v1 Announce Type: new Abstract: LLMs deployed in high-stakes domains face fundamental reliability challenges: hallucinations, inconsistencies, and privacy vulnerabilities introduce una

OmniToM: Benchmarking Theory of Mind in LLMs via Explicit Belief Modeling

Model ReleasesDGX agent

arXiv:2605.26322v1 Announce Type: new Abstract: Theory of Mind (ToM), the ability to infer others' knowledge, intentions, and emotions, is commonly evaluated in large language models (LLMs) using end-

On the Hidden Costs of Counterfactual Knowledge Training in LLM Unlearning

Model ReleasesDGX agent

arXiv:2605.27083v1 Announce Type: new Abstract: Counterfactual tuning (CFT) has emerged as a promising paradigm for Large Language Model (LLM) unlearning by training models to generate alternative fic

ORCA: An End-to-End Interactive Copilot for Optimized Root Cause Analysis

TutorialsDGX agent

arXiv:2605.27022v1 Announce Type: new Abstract: Causal analysis is a crucial task in many domains, including manufacturing, social science, and medicine. However, despite recent progress, the conceptu

PilotTTS: A Disciplined Modular Recipe for Competitive Speech Synthesis

Model ReleasesDGX agent

arXiv:2605.27258v1 Announce Type: cross Abstract: Building state-of-the-art text-to-speech (TTS) systems typically demands millions of hours of proprietary data and complex multi-stage architectures,

Position: Machine Learning for Heart Transplant Allocation Policy Optimization Should Account for Incentives

SafetyDGX agent

arXiv:2602.04990v3 Announce Type: replace Abstract: The allocation of scarce donor organs constitutes one of the most consequential algorithmic challenges in healthcare. While the field is rapidly tra

Qiskit QuantumKatas: Adapting Microsoft's Quantum Computing exercises for LLM evaluation

Model ReleasesDGX agent

arXiv:2605.27210v1 Announce Type: cross Abstract: We adapt Microsoft's QuantumKatas -- a well-established quantum computing curriculum -- from Q# to Qiskit, the most widely-adopted quantum computing f

ReasonOps: A Unified Operational Paradigm for Trustworthy Verified LLM Reasoning

SafetyDGX agent

arXiv:2605.27014v1 Announce Type: cross Abstract: Large Language Models (LLMs) have transformed artificial intelligence from primarily generative systems into increasingly capable reasoning agents. Re

SIA: Self Improving AI with Harness & Weight Updates

HardwareDGX agent

arXiv:2605.27276v1 Announce Type: new Abstract: Humans are the bottleneck in building and improving AI. Both the models and the agents that wrap them are written, tuned, and corrected by people. The l

SONAR-LLM: Autoregressive Transformer that Thinks in Sentence Embeddings and Speaks in Tokens

Model ReleasesDGX agent

arXiv:2508.05305v2 Announce Type: replace Abstract: The recently proposed Large Concept Model (LCM) generates text by predicting a sequence of sentence-level embeddings and training with either mean-s

sqlite AGENTS.md

AgentsDGX agent

sqlite AGENTS.md SQLite gained an AGENTS.md file five days ago - but it's not intended for their own development, it's presumably aimed at people who are pointing agents at the SQLite codebase. It inc

Stronger models do not always need lighter harnesses. Everyone believes more structured harnesses universally improve reliability, and that …

Model ReleasesDGX agent

Stronger models do not always need lighter harnesses. Everyone believes more structured harnesses universally improve reliability, and that higher-capability models need proportionally less structural

Synergetic Empowerment: Wireless Communications Meets Embodied Intelligence

AgentsDGX agent

arXiv:2509.10481v2 Announce Type: replace-cross Abstract: Wireless communication is evolving into an agent era, where large-scale agents with inherent embodied intelligence are not just users but acti

Telenor Nordics Customer Service self-help corpus

AgentsDGX agent

arXiv:2605.26891v1 Announce Type: new Abstract: This paper presents a multilingual customer service self-help corpus comprising 1,122 manually validated documents in Finnish, Danish, Norwegian, and Sw

the future of Continual Learning will rely on building to systems to ingest, understand, & apply knowledge from Agent Traces at scale had a …

AgentsDGX agent

the future of Continual Learning will rely on building to systems to ingest, understand, & apply knowledge from Agent Traces at scale had a blast presenting LangSmith Engine with @bentannyhill to show

The Necessity of a Unified Framework for LLM-Based Agent Evaluation

AgentsDGX agent

arXiv:2602.03238v2 Announce Type: replace Abstract: With the advent of Large Language Models (LLMs), general-purpose agents have seen fundamental advancements. However, evaluating these agents present

Towards Real-World Identification of Fatigued Muscle Groups via Musculoskeletal Simulation

ApplicationsDGX agent

arXiv:2605.26151v1 Announce Type: cross Abstract: Contactless diagnosis of musculoskeletal disorders can potentially improve population health as well as robot behaviours in collaborative settings. Ho

VISTA: An End-to-End Benchmark for Visual Spec-to-Web-App Coding Agents

Model ReleasesDGX agent

arXiv:2605.26144v1 Announce Type: cross Abstract: We present VISTA (VIsual Spec-To-App Benchmark), a benchmark for evaluating the end-to-end web-app generation capabilities of LLM-based agents. Unlike

26 May 2026

Active Learning for Stochastic Contextual Linear Bandits

SafetyDGX agent

arXiv:2605.24803v1 Announce Type: new Abstract: A key goal in stochastic contextual linear bandits is to efficiently learn a near-optimal policy. Prior algorithms for this problem learn a policy by st

Agent-X: Evaluating Deep Multimodal Reasoning in Vision-Centric Agentic Tasks

Model ReleasesDGX agent

arXiv:2505.24876v2 Announce Type: replace-cross Abstract: Deep reasoning is fundamental for solving complex tasks, especially in vision-centric scenarios that demand sequential, multimodal understandi

Authority Signals in Claude AI Health Citations: A Descriptive Analysis Using the Authority Signals Framework

Model ReleasesDGX agent

arXiv:2605.23921v1 Announce Type: cross Abstract: This study seeks to determine the authority signals used by Anthropic's Claude AI in its presentation of sources when answering consumer health questi

Beyond Literal Translation: Evaluating Cultural Effectiveness in Social Media UGC

Model ReleasesDGX agent

arXiv:2605.25626v1 Announce Type: new Abstract: Social media platforms enable large-scale cross-lingual communication, but translating user-generated content (UGC) remains challenging due to its infor

Bridging Earth and Space: A Survey on HAPS for Non-Terrestrial Networks

AgentsDGX agent

arXiv:2510.19731v3 Announce Type: replace-cross Abstract: HAPS are emerging as key enablers in the evolution of 6G wireless networks, bridging terrestrial and non-terrestrial infrastructures. Operatin

btw this will be the 3 year anniversary of the Rise of the AI Engineer blogpost. the industry keeps growing and growing, kinda scary to real…

ToolsDGX agent

btw this will be the 3 year anniversary of the Rise of the AI Engineer blogpost. the industry keeps growing and growing, kinda scary to realize i may never top it. AIE EU, AIE MIA and AIE SG reached 8

Building an Adversarial Malware Dataset by Family and Type: Generation, Evasion, and Poisoning Evaluation

Model ReleasesDGX agent

arXiv:2605.25937v1 Announce Type: cross Abstract: We present a dataset of adversarial malware samples derived from the public RawMal-TF collection of real-world malware binaries. Using a suite of adve

Can LLMs Time Travel? Enhancing Temporal Consistency in Legal Agentic Search through Reinforcement Learning

Model ReleasesDGX agent

arXiv:2605.25920v1 Announce Type: cross Abstract: While large language models (LLMs) augmented with agentic search capabilities show promise for legal reasoning, they overlook a fundamental constraint

Can LoRA Fusion Support Cross-Domain Tasks in Cloud-Edge Collaboration?

Model ReleasesDGX agent

arXiv:2605.23913v1 Announce Type: cross Abstract: Cloud-hosted large language models (LLMs) commonly rely on LoRA for domain adaptation, yet domain data are distributed across multiple edge devices an

Causal methods for LLM development and evaluation

SafetyDGX agent

arXiv:2605.25998v1 Announce Type: new Abstract: Large language model (LLM) development is currently driven by large-scale empirical iteration over data mixtures, reward models, routing strategies, and

Chain-of-Thought Hijacking

Model ReleasesDGX agent

arXiv:2510.26418v4 Announce Type: replace Abstract: Large Reasoning Models (LRMs) improve task performance through extended inference-time reasoning. Although previous studies suggest that longer reas

ChunkLLM: A Lightweight Pluggable Framework for Accelerating LLMs Inference

Model ReleasesDGX agent

arXiv:2510.02361v2 Announce Type: replace-cross Abstract: Transformer-based large models excel in natural language processing and computer vision, but face severe computational inefficiencies due to t

CITYREP: A Unified Benchmark for Urban Representations Across Cities, Tasks, and Modalities

Model ReleasesDGX agent

arXiv:2605.26036v1 Announce Type: new Abstract: Urban representation learning encodes complex urban environments into general-purpose embeddings for diverse downstream tasks and emerging urban foundat

ContextEcho: A Benchmark for Persona Drift in Long Agentic-Coding Sessions

Model ReleasesDGX agent

arXiv:2605.24279v1 Announce Type: new Abstract: A frontier language model's acknowledged 'helpful programming assistant' persona does not survive long agentic-coding sessions in the deployment regime

Demystifying the Mythos or Disrupting Bugonomics? From Zero-Day Asymmetry to Defender Remediation Throughput

TutorialsDGX agent

arXiv:2605.24632v1 Announce Type: cross Abstract: Recent demonstrations of large language models producing candidate and confirmed vulnerabilities in production software have renewed the narrative tha

Endometriosis 💛

IndustryDGX agent

Endometriosis is a chronic condition where tissue similar to the uterine lining grows outside the uterus, causing pain, inflammation, and potentially affecting fertility. The condition affects million

ERNIE-Image Technical Report

Model ReleasesDGX agent

arXiv:2605.25347v1 Announce Type: cross Abstract: We introduce ERNIE-Image, an open-source text-to-image generation model built upon an 8B single-stream DiT architecture. ERNIE-Image aims to bridge th

Extending Embodied Question Answering from Perception to Decision

Model ReleasesDGX agent

arXiv:2605.25813v1 Announce Type: new Abstract: Embodied Question Answering (EQA) connects perception, reasoning, and interaction within embodied environments. However, existing datasets and benchmark

FG-CLIP 2: A Bilingual Fine-grained Vision-Language Alignment Model

Model ReleasesDGX agent

arXiv:2510.10921v3 Announce Type: replace-cross Abstract: Fine-grained vision-language understanding requires precise alignment between visual content and linguistic descriptions, a capability that re

From Automation to Collaboration: Human-in-the-Loop Methods for Safe and Trustworthy NLP

SafetyDGX agent

arXiv:2605.25226v1 Announce Type: new Abstract: Large language models are widely deployed in high-stakes NLP tasks, yet risks such as bias, hallucination, adversarial vulnerability and unreliable gene

← Previous
1…398399400401402…428
Next →