AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
Human
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,851 results
23 Apr 2026

KoALa-Bench: Evaluating Large Audio Language Models on Korean Speech Understanding and Faithfulness

Model ReleasesDGX agent

arXiv:2604.19782v1 Announce Type: cross Abstract: Recent advances in large audio language models (LALMs) have enabled multilingual speech understanding. However, benchmarks for evaluating LALMs remain

KOCO-BENCH: Can Large Language Models Leverage Domain Knowledge in Software Development?

Model ReleasesDGX agent

arXiv:2601.13240v2 Announce Type: replace-cross Abstract: Large language models (LLMs) excel at general programming but struggle with domain-specific software development, necessitating domain special

Language-Coupled Reinforcement Learning for Multilingual Retrieval-Augmented Generation

SafetyDGX agent
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2601.14896v2 Announce Type: replace Abstract: Multilingual retrieval-augmented generation (MRAG) requires models to effectively acquire and integrate beneficial external knowledge from multiling

Language Models Learn Universal Representations of Numbers and Here's Why You Should Care

TutorialsDGX agent

arXiv:2510.26285v2 Announce Type: replace-cross Abstract: Prior work has shown that large language models (LLMs) often converge to accurate input embedding for numbers, based on sinusoidal representat

LaplacianFormer:Rethinking Linear Attention with Laplacian Kernel

HardwareDGX agent

arXiv:2604.20368v1 Announce Type: cross Abstract: The quadratic complexity of softmax attention presents a major obstacle for scaling Transformers to high-resolution vision tasks. Existing linear atte

Large Language Models Meet Biomedical Knowledge Graphs for Mechanistically Grounded Therapeutic Prioritization

Model ReleasesDGX agent

arXiv:2604.19815v1 Announce Type: new Abstract: Drug repurposing is often framed as a candidate identification task, but existing approaches provide limited guidance for distinguishing biologically pl

Large Language Models Outperform Humans in Fraud Detection and Resistance to Motivated Investor Pressure

Model ReleasesDGX agent

arXiv:2604.20652v1 Announce Type: new Abstract: Large language models trained on human feedback may suppress fraud warnings when investors arrive already persuaded of a fraudulent opportunity. We test

Large language models perceive cities through a culturally uneven baseline

SafetyDGX agent

arXiv:2604.20048v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used to describe, evaluate and interpret places, yet it remains unclear whether they do so from a cultural

Latent Stochastic Interpolants

Model ReleasesDGX agent

arXiv:2506.02276v2 Announce Type: replace Abstract: Stochastic Interpolants (SI) is a powerful framework for generative modeling, capable of flexibly transforming between two probability distributions

LayerTracer: A Joint Task-Particle and Vulnerable-Layer Analysis framework for Arbitrary Large Language Model Architectures

Model ReleasesDGX agent

arXiv:2604.20556v1 Announce Type: cross Abstract: Currently, Large Language Models (LLMs) feature a diversified architectural landscape, including traditional Transformer, GateDeltaNet, and Mamba. How

LEAD: Breaking the No-Recovery Bottleneck in Long-Horizon Reasoning

Local AiDGX agent

arXiv:2603.06870v2 Announce Type: replace Abstract: Long-horizon execution in Large Language Models (LLMs) remains unstable even when high-level strategies are provided. Evaluating on controlled algor

Learn2Synth: Learning Optimal Data Synthesis Using Hypergradients for Brain Image Segmentation

ApplicationsDGX agent

arXiv:2411.16719v4 Announce Type: replace Abstract: Domain randomization through synthesis is a powerful strategy to train networks that are unbiased with respect to the domain of the input images. Ra

Learning Multi-Modal Whole-Body Control for Real-World Humanoid Robots

ApplicationsDGX agent

arXiv:2408.07295v4 Announce Type: replace-cross Abstract: A major challenge in humanoid robotics is designing a unified interface for commanding diverse whole-body behaviors, from precise footstep seq

Learning Spatial-Temporal Coherent Correlations for Speech-Preserving Facial Expression Manipulation

Local AiDGX agent

arXiv:2604.20226v1 Announce Type: new Abstract: Speech-preserving facial expression manipulation (SPFEM) aims to modify facial emotions while meticulously maintaining the mouth animation associated wi

Learning to count small and clustered objects with application to bacterial colonies

SafetyDGX agent

arXiv:2604.20030v1 Announce Type: new Abstract: Automated bacterial colony counting from images is an important technique to obtain data required for the development of vaccines and antibiotics. Howev

Learning to Evolve: A Self-Improving Framework for Multi-Agent Systems via Textual Parameter Graph Optimization

Model ReleasesDGX agent

arXiv:2604.20714v1 Announce Type: new Abstract: Designing and optimizing multi-agent systems (MAS) is a complex, labor-intensive process of 'Agent Engineering.' Existing automatic optimization methods

Learning to Solve the Quadratic Assignment Problem with Warm-Started MCMC Finetuning

ApplicationsDGX agent

arXiv:2604.20109v1 Announce Type: cross Abstract: The quadratic assignment problem (QAP) is a fundamental NP-hard task that poses significant challenges for both traditional heuristics and modern lear

Learning When Not to Decide: A Framework for Overcoming Factual Presumptuousness in AI Adjudication

Model ReleasesDGX agent

arXiv:2604.19895v1 Announce Type: new Abstract: A well-known limitation of AI systems is presumptuousness: the tendency of AI systems to provide confident answers when information may be lacking. This

Less Languages, Less Tokens: An Efficient Unified Logic Cross-lingual Chain-of-Thought Reasoning Framework

Model ReleasesDGX agent

arXiv:2604.20090v1 Announce Type: new Abstract: Cross-lingual chain-of-thought (XCoT) with self-consistency markedly enhances multilingual reasoning, yet existing methods remain costly due to extensiv

Lever: Inference-Time Policy Reuse under Support Constraints

SafetyDGX agent

arXiv:2604.20174v1 Announce Type: new Abstract: Reinforcement learning (RL) policies are typically trained for fixed objectives, making reuse difficult when task requirements change. We study inferenc

Lexicographic Minimum-Violation Motion Planning using Signal Temporal Logic

AgentsDGX agent

arXiv:2604.20428v1 Announce Type: new Abstract: Motion planning for autonomous vehicles often requires satisfying multiple conditionally conflicting specifications. In situations where not all specifi

LEXIS: LatEnt ProXimal Interaction Signatures for 3D HOI from an Image

ResearchDGX agent

arXiv:2604.20800v1 Announce Type: new Abstract: Reconstructing 3D Human-Object Interaction from an RGB image is essential for perceptive systems. Yet, this remains challenging as it requires capturing

Lifecycle-Aware Federated Continual Learning in Mobile Autonomous Systems

AgentsDGX agent

arXiv:2604.20745v1 Announce Type: cross Abstract: Federated continual learning (FCL) allows distributed autonomous fleets to adapt collaboratively to evolving terrain types across extended mission lif

Lightweight LLM Agent Memory with Small Language Models

AgentsDGX agent

arXiv:2604.07798v3 Announce Type: replace Abstract: Although LLM agents can leverage tools for complex tasks, they still need memory to maintain cross-turn consistency and accumulate reusable informat

LiteResearcher: A Scalable Agentic RL Training Framework for Deep Research Agent

Model ReleasesDGX agent

arXiv:2604.17931v2 Announce Type: replace Abstract: Reinforcement Learning (RL) has emerged as a powerful training paradigm for LLM-based agents. However, scaling agentic RL for deep research remains

LLaDA2.0-Uni: Unifying Multimodal Understanding and Generation with Diffusion Large Language Model

ResearchDGX agent

arXiv:2604.20796v1 Announce Type: new Abstract: We present LLaDA2.0-Uni, a unified discrete diffusion large language model (dLLM) that supports multimodal understanding and generation within a nativel

LLAMADRS: Evaluating Open-Source LLMs on Real Clinical Interviews--To Reason or Not to Reason?

Model ReleasesDGX agent

arXiv:2501.03624v2 Announce Type: replace-cross Abstract: Large language models (LLMs) excel on many NLP benchmarks, but their behavior on real-world, semi-structured prediction remains underexplored.

LLM Agents Grounded in Self-Reports Enable General-Purpose Simulation of Individuals

AgentsDGX agent

arXiv:2411.10109v2 Announce Type: replace Abstract: Machine learning can predict human behavior well when substantial structured data and well-defined outcomes are available, but these models are typi

LLM Agents Predict Social Media Reactions but Do Not Outperform Text Classifiers: Benchmarking Simulation Accuracy Using 120K+ Personas of 1511 Humans

SafetyDGX agent

arXiv:2604.19787v1 Announce Type: cross Abstract: Social media platforms mediate how billions form opinions and engage with public discourse. As autonomous AI agents increasingly participate in these

LLM-guided phase diagram construction through high-throughput experimentation

Model ReleasesDGX agent

arXiv:2604.20304v1 Announce Type: cross Abstract: Constructing phase diagrams for multicomponent alloys requires extensive experimental measurements and is a time-consuming task. Here we investigate w

LLM-Guided Safety Agent for Edge Robotics with an ISO-Compliant Perception-Compute-Control Architecture

SafetyDGX agent

arXiv:2604.20193v1 Announce Type: new Abstract: Ensuring functional safety in human-robot interaction is challenging because AI perception is inherently probabilistic, whereas industrial standards req

LLM StructCore: Schema-Guided Reasoning Condensation and Deterministic Compilation

ResearchDGX agent

arXiv:2604.20560v1 Announce Type: new Abstract: Automatically filling Case Report Forms (CRFs) from clinical notes is challenging due to noisy language, strict output contracts, and the high cost of f

LLMs Can Get 'Brain Rot': A Pilot Study on Twitter/X

SafetyDGX agent

arXiv:2510.13928v2 Announce Type: replace-cross Abstract: We propose and test the LLM Brain Rot Hypothesis: continual exposure to junk web text induces lasting cognitive decline in large language mode

Local Diffusion Models and Phases of Data Distributions

TutorialsDGX agent

arXiv:2508.06614v2 Announce Type: replace Abstract: As a class of generative artificial intelligence frameworks inspired by statistical physics, diffusion models have shown extraordinary performance i

Locate-Then-Examine: Grounded Region Reasoning Improves Detection of AI-Generated Images

Local AiDGX agent

arXiv:2510.04225v2 Announce Type: replace-cross Abstract: The rapid growth of AI-generated imagery has blurred the boundary between real and synthetic content, raising practical concerns for digital i

Location-Aware Pretraining for Medical Difference Visual Question Answering

ResearchDGX agent

arXiv:2603.04950v2 Announce Type: replace-cross Abstract: Differential medical VQA models compare multiple images to identify clinically meaningful changes and rely on vision encoders to capture fine-

LoRA-FA: Efficient and Effective Low Rank Representation Fine-tuning

Model ReleasesDGX agent

arXiv:2308.03303v2 Announce Type: replace Abstract: Fine-tuning large language models (LLMs) is crucial for improving their performance on downstream tasks, but full-parameter fine-tuning (Full-FT) is

Lucky High Dynamic Range Smartphone Imaging

ResearchDGX agent

arXiv:2604.19976v1 Announce Type: new Abstract: While the human eye can perceive an impressive twenty stops of dynamic range, smartphone camera sensors remain limited to about twelve stops despite dec

Machine Learning for Two-Stage Graph Sparsification for the Travelling Salesman Problem

TutorialsDGX agent

arXiv:2604.20236v1 Announce Type: new Abstract: High-performance TSP solvers like LKH search within a sparsified candidate graph rather than over all possible edges. Graph sparsification is non-trivia

Machine learning moment closure models for the radiative transfer equation IV: enforcing symmetrizable hyperbolicity in two dimensions

ResearchDGX agent

arXiv:2604.20143v1 Announce Type: cross Abstract: This is our fourth work in the series on machine learning (ML) moment closure models for the radiative transfer equation (RTE). In the first three pap

MambaLiteUNet: Cross-Gated Adaptive Feature Fusion for Robust Skin Lesion Segmentation

Model ReleasesDGX agent

arXiv:2604.20286v1 Announce Type: cross Abstract: Recent segmentation models have demonstrated promising efficiency by aggressively reducing parameter counts and computational complexity. However, the

MAPRPose: Mask-Aware Proposal and Amodal Refinement for Multi-Object 6D Pose Estimation

Model ReleasesDGX agent

arXiv:2604.20650v1 Announce Type: new Abstract: 6D object pose estimation in cluttered scenes remains challenging due to severe occlusion and sensor noise. We propose MAPRPose, a two-stage framework t

Markov reads Pushkin, again: A statistical journey into the poetic world of Evgenij Onegin

ResearchDGX agent

arXiv:2604.20221v1 Announce Type: new Abstract: This study applies symbolic time series analysis and Markov modeling to explore the phonological structure of Evgenij Onegin-as captured through a graph

MasconCube: Fast and Accurate Gravity Modeling with an Explicit Representation

ResearchDGX agent

arXiv:2509.08607v3 Announce Type: replace-cross Abstract: The geodesy of irregularly shaped small bodies presents fundamental challenges for gravitational field modeling, particularly as deep space ex

MATT-Diff: Multimodal Active Target Tracking by Diffusion Policy

SafetyDGX agent

arXiv:2511.11931v2 Announce Type: replace Abstract: This paper proposes MATT-Diff: Multimodal Active Target Tracking by Diffusion Policy, a control policy for active multi-target tracking using a mobi

Maximum Entropy Semi-Supervised Inverse Reinforcement Learning

ResearchDGX agent

arXiv:2604.20074v1 Announce Type: new Abstract: A popular approach to apprenticeship learning (AL) is to formulate it as an inverse reinforcement learning (IRL) problem. The MaxEnt-IRL algorithm succe

Maximum Likelihood Reconstruction for Multi-Look Digital Holography with Markov-Modeled Speckle Correlation

ResearchDGX agent

arXiv:2604.20154v1 Announce Type: cross Abstract: Multi-look acquisition is a widely used strategy for reducing speckle noise in coherent imaging systems such as digital holography. By acquiring multi

MD-Face: MoE-Enhanced Label-Free Disentangled Representation for Interactive Facial Attribute Editing

TutorialsDGX agent

arXiv:2604.20317v1 Announce Type: new Abstract: GAN-based facial attribute editing is widely used in virtual avatars and social media but often suffers from attribute entanglement, where modifying one

Measuring Creativity in the Age of Generative AI: Distinguishing Human and AI-Generated Creative Performance in Hiring and Talent Systems

ResearchDGX agent

arXiv:2604.19799v1 Announce Type: cross Abstract: Generative AI is rapidly transforming how organizations create value and evaluate talent. While large language models enhance baseline output quality,

Measuring the Machine: Evaluating Generative AI as Pluralist Sociotechical Systems

Model ReleasesDGX agent

arXiv:2604.20545v1 Announce Type: new Abstract: In measurement theory, instruments do not simply record reality; they help constitute what is observed. The same holds for generative AI evaluation: ben

Mechanistic Interpretability of Large-Scale Counting in LLMs through a System-2 Strategy

ResearchDGX agent

arXiv:2601.02989v2 Announce Type: replace Abstract: Large language models (LLMs), despite strong performance on complex mathematical problems, exhibit systematic limitations in counting tasks. This is

Mechanistic Interpretability Tool for AI Weather Models

ResearchDGX agent

arXiv:2604.20467v1 Announce Type: cross Abstract: Artificial Intelligence (AI) weather models are improving rapidly, and their forecasts are already competitive with long-established traditional Numer

MedSkillAudit: A Domain-Specific Audit Framework for Medical Research Agent Skills

SafetyDGX agent

arXiv:2604.20441v1 Announce Type: new Abstract: Background: Agent skills are increasingly deployed as modular, reusable capability units in AI agent systems. Medical research agent skills require safe

Membership Inference for Contrastive Pre-training Models with Text-only PII Queries

SafetyDGX agent

arXiv:2603.14222v2 Announce Type: replace-cross Abstract: Contrastive pretraining models such as CLIP and CLAP, serve as the ubiquitous perceptual backbones for modern multimodal large models, yet the

Memorization, Emergence, and Explaining Reversal Failures: A Controlled Study of Relational Semantics in LLMs

SafetyDGX agent

arXiv:2601.02931v2 Announce Type: replace Abstract: Autoregressive LLMs perform well on relational tasks that require linking entities via relational words (e.g., father/son, friend), but it is unclea

Memory-Augmented LLM-based Multi-Agent System for Automated Feature Generation on Tabular Data

AgentsDGX agent

arXiv:2604.20261v1 Announce Type: new Abstract: Automated feature generation extracts informative features from raw tabular data without manual intervention and is crucial for accurate, generalizable

Meta Additive Model: Interpretable Sparse Learning With Auto Weighting

ApplicationsDGX agent

arXiv:2604.20111v1 Announce Type: cross Abstract: Sparse additive models have attracted much attention in high-dimensional data analysis due to their flexible representation and strong interpretabilit

Meta-Tool: Efficient Few-Shot Tool Adaptation for Small Language Models

Model ReleasesDGX agent

arXiv:2604.20148v1 Announce Type: cross Abstract: Can small language models achieve strong tool-use performance without complex adaptation mechanisms? This paper investigates this question through Met

MetaboNet: The Largest Publicly Available Consolidated Dataset for Type 1 Diabetes Management

Model ReleasesDGX agent

arXiv:2601.11505v2 Announce Type: replace-cross Abstract: Progress in Type 1 Diabetes (T1D) algorithm development is limited by the fragmentation and lack of standardization across existing T1D manage

MGDA-Decoupled: Geometry-Aware Multi-Objective Optimisation for DPO-based LLM Alignment

SafetyDGX agent

arXiv:2604.20685v1 Announce Type: new Abstract: Aligning large language models (LLMs) to desirable human values requires balancing multiple, potentially conflicting objectives such as helpfulness, tru

← Previous
1…863864865866867…998
Next →