AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
Human
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,851 results
24 Apr 2026

Language as a Latent Variable for Reasoning Optimization

Model ReleasesDGX agent

arXiv:2604.21593v1 Announce Type: new Abstract: As LLMs reduce English-centric bias, a surprising trend emerges: non-English responses sometimes outperform English on reasoning tasks. We hypothesize t

Language-Conditioned Safe Trajectory Generation for Spacecraft Rendezvous

SafetyDGX agent

arXiv:2512.09111v3 Announce Type: replace-cross Abstract: Reliable real-time trajectory generation is essential for future autonomous spacecraft. While recent progress in nonconvex guidance and contro

Large-Scale Data Parallelization of Product Quantization and Inverted Indexing Using Dask

ResearchDGX agent

arXiv:2604.21645v1 Announce Type: new Abstract: Large-scale Nearest Neighbor (NN) search, though widely utilized in the similarity search field, remains challenged by the computational limitations inh

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Latent Denoising Improves Visual Alignment in Large Multimodal Models

Model ReleasesDGX agent

arXiv:2604.21343v1 Announce Type: new Abstract: Large Multimodal Models (LMMs) such as LLaVA are typically trained with an autoregressive language modeling objective, providing only indirect supervisi

LatRef-Diff: Latent and Reference-Guided Diffusion for Facial Attribute Editing and Style Manipulation

ResearchDGX agent

arXiv:2604.21279v1 Announce Type: new Abstract: Facial attribute editing and style manipulation are crucial for applications like virtual avatars and photo editing. However, achieving precise control

Learn Weightlessness: Imitate Non-Self-Stabilizing Motions on Humanoid Robot

TutorialsDGX agent

arXiv:2604.21351v1 Announce Type: new Abstract: The integration of imitation and reinforcement learning has enabled remarkable advances in humanoid whole-body control, facilitating diverse human-like

Learning Dynamic Representations and Policies from Multimodal Clinical Time-Series with Informative Missingness

SafetyDGX agent

arXiv:2604.21235v1 Announce Type: cross Abstract: Multimodal clinical records contain structured measurements and clinical notes recorded over time, offering rich temporal information about the evolut

Learning Linear Regression with Low-Rank Tasks in-Context

TutorialsDGX agent

arXiv:2510.04548v2 Announce Type: replace-cross Abstract: In-context learning (ICL) is a key building block of modern large language models, yet its theoretical mechanisms remain poorly understood. It

Learning Physics from Pretrained Video Models: A Multimodal Continuous and Sequential World Interaction Models for Robotic Manipulation

SafetyDGX agent

arXiv:2603.00110v2 Announce Type: replace Abstract: The scarcity of large-scale robotic data has motivated the repurposing of foundation models from other modalities for policy learning. In this work,

Learning Reasoning Reward Models from Expert Demonstration via Inverse Reinforcement Learning

Local AiDGX agent

arXiv:2510.01857v3 Announce Type: replace Abstract: Current approaches to improving reasoning in large language models (LLMs) primarily rely on either supervised fine-tuning (SFT) over expert traces o

Learning State-Tracking from Code Using Linear RNNs

ResearchDGX agent

arXiv:2602.14814v2 Announce Type: replace-cross Abstract: Over the last years, state-tracking tasks, particularly permutation composition, have become a testbed to understand the limits of sequence mo

Learning to Communicate: Toward End-to-End Optimization of Multi-Agent Language Systems

Model ReleasesDGX agent

arXiv:2604.21794v1 Announce Type: new Abstract: Multi-agent systems built on large language models have shown strong performance on complex reasoning tasks, yet most work focuses on agent roles and or

Learning to Emulate Chaos: Adversarial Optimal Transport Regularization

TutorialsDGX agent

arXiv:2604.21097v1 Announce Type: cross Abstract: Chaos arises in many complex dynamical systems, from weather to power grids, but is difficult to accurately model using data-driven emulators, includi

Leveraging Multimodal LLMs for Built Environment and Housing Attribute Assessment from Street-View Imagery

Model ReleasesDGX agent

arXiv:2604.21102v1 Announce Type: cross Abstract: We present a novel framework for automatically evaluating building conditions nationwide in the United States by leveraging large language models (LLM

Linear Image Generation by Synthesizing Exposure Brackets

ResearchDGX agent

arXiv:2604.21008v1 Announce Type: new Abstract: The life of a photo begins with photons striking the sensor, whose signals are passed through a sophisticated image signal processing (ISP) pipeline to

Listen and Chant Before You Read: The Ladder of Beauty in LM Pre-Training

ResearchDGX agent

arXiv:2604.21265v1 Announce Type: new Abstract: We show that pre-training a Transformer on music before language significantly accelerates language acquisition. Using piano performances (MAESTRO datas

LiveSense: A Real-Time Wi-Fi Sensing Platform for Range-Doppler on COTS Laptop

Local AiDGX agent

arXiv:2603.06545v2 Announce Type: replace-cross Abstract: We present LiveSense - a cross-platform that transforms a commercial off-the-shelf (COTS) Wi-Fi Network Interface Card (NIC) on a laptop into

LiveVLM: Efficient Online Video Understanding via Streaming-Oriented KV Cache and Retrieval

Model ReleasesDGX agent

arXiv:2505.15269v2 Announce Type: replace Abstract: Recent developments in Video Large Language Models (Video LLMs) have enabled models to process hour-long videos and exhibit exceptional performance.

Local Neighborhood Instability in Parametric Projections: Quantitative and Visual Analysis

SafetyDGX agent

arXiv:2604.21617v1 Announce Type: new Abstract: Parametric projections let analysts embed new points in real time, but input variations from measurement noise or data drift can produce unpredictable s

Locating acts of mechanistic reasoning in student team conversations with mechanistic machine learning

SafetyDGX agent

arXiv:2604.21870v1 Announce Type: cross Abstract: STEM education researchers are often interested in identifying moments of students' mechanistic reasoning for deeper analysis, but have limited capaci

Logic Jailbreak: Efficiently Unlocking LLM Safety Restrictions Through Formal Logical Expression

SafetyDGX agent

arXiv:2505.13527v3 Announce Type: replace-cross Abstract: Despite substantial advancements in aligning large language models (LLMs) with human values, current safety mechanisms remain susceptible to j

Long-Horizon Manipulation via Trace-Conditioned VLA Planning

ResearchDGX agent

arXiv:2604.21924v1 Announce Type: new Abstract: Long-horizon manipulation remains challenging for vision-language-action (VLA) policies: real tasks are multi-step, progress-dependent, and brittle to c

Losing our Tail, Again: (Un)Natural Selection & Multilingual LLMs

ResearchDGX agent

arXiv:2507.03933v3 Announce Type: replace Abstract: Multilingual Large Language Models considerably changed how technologies influence language. While previous technologies could mediate or assist hum

Low Cost, High Efficiency: LiDAR Place Recognition in Vineyards with Matryoshka Representation Learning

ResearchDGX agent

arXiv:2601.18714v2 Announce Type: replace-cross Abstract: Localization in agricultural environments is challenging due to their unstructured nature and lack of distinctive landmarks. Although agricult

Low-Rank Adaptation Redux for Large Models

Model ReleasesDGX agent

arXiv:2604.21905v1 Announce Type: new Abstract: Low-rank adaptation (LoRA) has emerged as the de facto standard for parameter-efficient fine-tuning (PEFT) of foundation models, enabling the adaptation

LRDUN: A Low-Rank Deep Unfolding Network for Efficient Spectral Compressive Imaging

ResearchDGX agent

arXiv:2511.18513v2 Announce Type: replace Abstract: Deep unfolding networks (DUNs) have achieved remarkable success and become the mainstream paradigm for spectral compressive imaging (SCI) reconstruc

M-CARE: Standardized Clinical Case Reporting for AI Model Behavioral Disorders, with a 20-Case Atlas and Experimental Validation

ResearchDGX agent

arXiv:2604.20871v1 Announce Type: cross Abstract: We introduce M-CARE (Model Clinical Assessment and Reporting for Evaluation), a clinical case report framework for AI model behavioral disorders adapt

Machine Behavior in Relational Moral Dilemmas: Moral Rightness, Predicted Human Behavior, and Model Decisions

SafetyDGX agent

arXiv:2604.21871v1 Announce Type: new Abstract: Human moral judgment is context-dependent and modulated by interpersonal relationships. As large language models (LLMs) increasingly function as decisio

Machine learning and digital pragmatics: Which word category influences emoji use most?

ResearchDGX agent

arXiv:2604.21108v1 Announce Type: new Abstract: This study investigates Machine Learning (ML) in the prediction of emojis in Arabic tweets employing the (state-of-the-art) MARBERT model. A corpus of 1

Mapping the Political Discourse in the Brazilian Chamber of Deputies: A Multi-Faceted Computational Approach

ApplicationsDGX agent

arXiv:2604.21897v1 Announce Type: new Abstract: Analyses of legislative behavior often rely on voting records, overlooking the rich semantic and rhetorical content of political speech. In this paper,

MaskDiME: Adaptive Masked Diffusion for Precise and Efficient Visual Counterfactual Explanations

Model ReleasesDGX agent

arXiv:2602.18792v3 Announce Type: replace Abstract: Visual counterfactual explanations aim to reveal the minimal semantic modifications that can alter a model's prediction, providing causal and interp

Materialistic RIR: Material Conditioned Realistic RIR Generation

ResearchDGX agent

arXiv:2604.21119v1 Announce Type: cross Abstract: Rings like gold, thuds like wood! The sound we hear in a scene is shaped not only by the spatial layout of the environment but also by the materials o

MathDuels: Evaluating LLMs as Problem Posers and Solvers

Model ReleasesDGX agent

arXiv:2604.21916v1 Announce Type: new Abstract: As frontier language models attain near-ceiling performance on static mathematical benchmarks, existing evaluations are increasingly unable to different

MATRAG: Multi-Agent Transparent Retrieval-Augmented Generation for Explainable Recommendations

Model ReleasesDGX agent

arXiv:2604.20848v1 Announce Type: cross Abstract: Large Language Model (LLM)-based recommendation systems have demonstrated remarkable capabilities in understanding user preferences and generating per

MCAP: Deployment-Time Layer Profiling for Memory-Constrained LLM Inference

Model ReleasesDGX agent

arXiv:2604.21026v1 Announce Type: new Abstract: Deploying large language models to heterogeneous hardware is often constrained by memory, not compute. We introduce MCAP (Monte Carlo Activation Profili

mcdok at SemEval-2026 Task 13: Finetuning LLMs for Detection of Machine-Generated Code

ResearchDGX agent

arXiv:2604.21365v1 Announce Type: cross Abstract: Multi-domain detection of the machine-generated code snippets in various programming languages is a challenging task. SemEval-2026 Task~13 copes with

Measure Twice, Click Once: Co-evolving Proposer and Visual Critic via Reinforcement Learning for GUI Grounding

Local AiDGX agent

arXiv:2604.21268v1 Announce Type: cross Abstract: Graphical User Interface (GUI) grounding requires mapping natural language instructions to precise pixel coordinates. However, due to visually homogen

Measuring and Exploiting Contextual Bias in LLM-Assisted Security Code Review

SafetyDGX agent

arXiv:2603.18740v2 Announce Type: replace-cross Abstract: Automated Code Review (ACR) systems integrating Large Language Models (LLMs) are increasingly adopted in software development workflows, rangi

Measuring Opinion Bias and Sycophancy via LLM-based Coercion

Model ReleasesDGX agent

arXiv:2604.21564v1 Announce Type: new Abstract: Large language models increasingly shape the information people consume: they are embedded in search, consulted for professional advice, deployed as age

mGRADE: Minimal Recurrent Gating Meets Delay Convolutions for Lightweight Sequence Modeling

Model ReleasesDGX agent

arXiv:2507.01829v2 Announce Type: replace-cross Abstract: Multi-timescale sequence modeling relies on capturing both local fast dynamics and global slow context; yet, maintaining these capabilities un

Micro-DualNet: Dual-Path Spatio-Temporal Network for Micro-Action Recognition

ResearchDGX agent

arXiv:2604.21011v1 Announce Type: new Abstract: Micro-actions are subtle, localized movements lasting 1-3 seconds such as scratching one's head or tapping fingers. Such subtle actions are essential fo

MiMIC: Mitigating Visual Modality Collapse in Universal Multimodal Retrieval While Avoiding Semantic Misalignment

ResearchDGX agent

arXiv:2604.21326v1 Announce Type: cross Abstract: Universal Multimodal Retrieval (UMR) aims to map different modalities (e.g., visual and textual) into a shared embedding space for multi-modal retriev

Mind the Gap: Optimal and Equitable Encouragement Policies

Local AiDGX agent

arXiv:2309.07176v5 Announce Type: replace Abstract: In consequential domains, it is often impossible to compel individuals to take treatment, so that optimal policy rules are merely suggestions in the

Mind the Prompt: Self-adaptive Generation of Task Plan Explanations via LLMs

SafetyDGX agent

arXiv:2604.21092v1 Announce Type: new Abstract: Integrating Large Language Models (LLMs) into complex software systems enables the generation of human-understandable explanations of opaque AI processe

Misinformation Span Detection in Videos via Audio Transcripts

ResearchDGX agent

arXiv:2604.21767v1 Announce Type: new Abstract: Online misinformation is one of the most challenging issues lately, yielding severe consequences, including political polarization, attacks on democracy

MISTY: High-Throughput Motion Planning via Mixer-based Single-step Drifting

Model ReleasesDGX agent

arXiv:2604.21489v1 Announce Type: cross Abstract: Multi-modal trajectory generation is essential for safe autonomous driving, yet existing diffusion-based planners suffer from high inference latency d

Mitigating Lost in Multi-turn Conversation via Curriculum RL with Verifiable Accuracy and Abstention Rewards

SafetyDGX agent

arXiv:2510.18731v2 Announce Type: replace-cross Abstract: Large Language Models demonstrate strong capabilities in single-turn instruction following but suffer from Lost-in-Conversation (LiC), a degra

Mixture of Sequence: Theme-Aware Mixture-of-Experts for Long-Sequence Recommendation

TutorialsDGX agent

arXiv:2604.20858v1 Announce Type: cross Abstract: Sequential recommendation has rapidly advanced in click-through rate prediction due to its ability to model dynamic user interests. A key challenge, h

MKJ at SemEval-2026 Task 9: A Comparative Study of Generalist, Specialist, and Ensemble Strategies for Multilingual Polarization

ResearchDGX agent

arXiv:2604.21370v1 Announce Type: new Abstract: We present a systematic study of multilingual polarization detection across 22 languages for SemEval-2026 Task 9 (Subtask 1), contrasting multilingual g

Modulating Cross-Modal Convergence with Single-Stimulus, Intra-Modal Dispersion

SafetyDGX agent

arXiv:2604.21836v1 Announce Type: cross Abstract: Neural networks exhibit a remarkable degree of representational convergence across diverse architectures, training objectives, and even data modalitie

Multi-Agent Empowerment and Emergence of Complex Behavior in Groups

AgentsDGX agent

arXiv:2604.21155v1 Announce Type: new Abstract: Intrinsic motivations are receiving increasing attention, i.e. behavioral incentives that are not engineered, but emerge from the interaction of an agen

Multilingual and Domain-Agnostic Tip-of-the-Tongue Query Generation for Simulated Evaluation

Model ReleasesDGX agent

arXiv:2604.21096v1 Announce Type: cross Abstract: Tip-of-the-Tongue (ToT) retrieval benchmarks have largely focused on English, limiting their applicability to multilingual information access. In this

Multilinguality at the Edge: Developing Language Models for the Global South

ResearchDGX agent

arXiv:2604.21637v1 Announce Type: new Abstract: Where and how language models (LMs) are deployed determines who can benefit from them. However, there are several challenges that prevent effective depl

Multimodal Bayesian Network for Robust Assessment of Casualties in Autonomous Triage

AgentsDGX agent

arXiv:2512.18908v2 Announce Type: replace Abstract: Mass Casualty Incidents can overwhelm emergency medical systems and resulting delays or errors in the assessment of casualties can lead to preventab

Multimodal Protein Language Models for Enzyme Kinetic Parameters: From Substrate Recognition to Conformational Adaptation

SafetyDGX agent

arXiv:2603.12845v2 Announce Type: replace Abstract: Predicting enzyme kinetic parameters quantifies how efficiently an enzyme catalyzes a specific substrate under defined biochemical conditions. Canon

Multiscale Super Resolution without Image Priors

ResearchDGX agent

arXiv:2604.21810v1 Announce Type: new Abstract: We address the ambiguities in the super-resolution problem under translation. We demonstrate that combinations of low-resolution images at different sca

Musical Score Understanding Benchmark: Evaluating Large Language Models' Comprehension of Complete Musical Scores

Model ReleasesDGX agent

arXiv:2511.20697v4 Announce Type: replace-cross Abstract: Understanding complete musical scores entails integrated reasoning over pitch, rhythm, harmony, and large-scale structure, yet the ability of

Navigating the Clutter: Waypoint-Based Bi-Level Planning for Multi-Robot Systems

Model ReleasesDGX agent

arXiv:2604.21138v1 Announce Type: cross Abstract: Multi-robot control in cluttered environments is a challenging problem that involves complex physical constraints, including robot-robot collisions, r

Nemobot Games: Crafting Strategic AI Gaming Agents for Interactive Learning with Large Language Models

Model ReleasesDGX agent

arXiv:2604.21896v1 Announce Type: new Abstract: This paper introduces a new paradigm for AI game programming, leveraging large language models (LLMs) to extend and operationalize Claude Shannon's taxo

Neural surrogates for crystal growth dynamics with variable supersaturation: explicit vs. implicit conditioning

Model ReleasesDGX agent

arXiv:2604.21753v1 Announce Type: cross Abstract: Simulations of crystal growth are performed by using Convolutional Recurrent Neural Network surrogate models, trained on a dataset of time sequences c

← Previous
1…853854855856857…998
Next →