AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
86,428Total entries
1Added by human
86,427Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
50,764 results
Model Releases

Knowledge Boundary Probing and Demand-Guided Intervention for LLM-Based Power System Code Generation

DGX agent

arXiv:2605.31478v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used to automate power-system analysis, but many utilities and energy-research labs require on-premise s

model-releasesarxiv-cs-cl
1 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Object-Informed Model Predictive Path Integral Control for Non-Prehensile Robot Manipulation

DGX agent

arXiv:2605.30778v1 Announce Type: new Abstract: Long-horizon planning for non-prehensile robot manipulation is challenging due to underactuated and discontinuous interactions. We propose a hierarchica

researcharxiv-cs-ro
1 Jun 2026
Model Releases

On the Robustness of Multilingual Text Embedding Rankings Across Learning Tasks, Languages, and Benchmark Datasets

DGX agent

arXiv:2605.31142v1 Announce Type: cross Abstract: Large-scale multilingual text embedding models play crucial role in both research and industry, yet their behavior in language-specific, multi-task se

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Query-focused and Memory-aware Reranker for Long Context Processing

DGX agent

arXiv:2602.12192v3 Announce Type: replace Abstract: Built upon the existing analysis of retrieval heads in large language models, we propose an alternative reranking framework that trains models to es

model-releasesarxiv-cs-cl
1 Jun 2026
Local Ai

Rethinking Efficient Crack Segmentation with Task-Aligned Structural-Directional Modeling

DGX agent

arXiv:2605.31048v1 Announce Type: new Abstract: Recent crack segmentation methods often follow generic semantic segmentation designs, using stronger backbones, hybrid CNN-Transformer-Mamba encoders, a

local-aiarxiv-cs-cv
1 Jun 2026
Safety

Scaling Multi-Agent Environment Co-Design with Diffusion Models

DGX agent

arXiv:2511.03100v2 Announce Type: replace-cross Abstract: The agent-environment co-design paradigm jointly optimises agent policies and environment configurations in search of improved system performa

safetyarxiv-cs-ai
1 Jun 2026
Model Releases

Steering LLMs? Actually, Sparse Autoencoders can outperform simple baselines

DGX agent

arXiv:2605.31183v1 Announce Type: cross Abstract: Sparse Autoencoders (SAEs) have been seen as a promising avenue for exploring the internals of Large Language Models (LLMs) and for steering model out

model-releasesarxiv-cs-ai
1 Jun 2026
Applications

VLM-GLoc: Vision-Language Model Enhanced Monte Carlo Localization for Robust Semantic Global Localization in Cluttered Quasi-Static Environments

DGX agent

arXiv:2605.30506v1 Announce Type: cross Abstract: Global localization in geometrically aliased, quasi-static environments such as grocery stores, offices, schools, and hospitals poses a significant ch

applicationsarxiv-cs-cv
1 Jun 2026
Model Releases

A Dual-Path Architecture for Scaling Compute and Capacity in LLMs

DGX agent

arXiv:2605.30202v1 Announce Type: new Abstract: Looped transformers apply a shared block multiple times and have emerged as a parameter-efficient route to scaling compute in language models. However,

model-releasesarxiv-cs-cl
29 May 2026
Applications

A Matter of Interest: Understanding Interestingness of Math Problems in Humans and Language Models

DGX agent

arXiv:2511.08548v2 Announce Type: replace Abstract: The evolution of mathematics is shaped importantly by interestingness: researchers choose which problems to pursue, and students choose which proble

applicationsarxiv-cs-ai
29 May 2026
Research

AnyMo: Scaling Any-Modality Conditional Motion Generation with Masked Modeling

DGX agent

arXiv:2605.29488v1 Announce Type: cross Abstract: Conditional human motion generation remains a fundamental challenge in computer vision and robotics. Despite significant progress, current methods are

researcharxiv-cs-ai
29 May 2026
Model Releases

AttuneBench: A Conversation-Based Benchmark for LLM Emotional Intelligence

DGX agent

arXiv:2605.21739v2 Announce Type: replace Abstract: Emotional intelligence (EI), the ability to perceive, understand, and respond appropriately to others' emotional states, is central to human communi

model-releasesarxiv-cs-ai
29 May 2026
Agents

BitTP: The Lightweight Trajectory Prediction Model with BitLLM for Edge-Devices

DGX agent

arXiv:2605.29705v1 Announce Type: new Abstract: Trajectory prediction is a fundamental task for autonomous systems, requiring complex reasoning about multi-agent interactions and intents. Large langua

agentsarxiv-cs-ai
29 May 2026
Model Releases

Dissecting the Black Box: Circuit-Level Analysis of LLM Vulnerability Detection

DGX agent

arXiv:2605.29901v1 Announce Type: cross Abstract: Large language models (LLMs) can detect software vulnerabilities, but how do they actually identify vulnerable code? We address this question using me

model-releasesarxiv-cs-lg
29 May 2026
Model Releases

DMC-CF: Dynamic Multimodal CounterFactual QA benchmark for Causal Reasoning

DGX agent

arXiv:2605.29339v1 Announce Type: new Abstract: With the rapid advancement of multimodal large language models (MLLMs), models have demonstrated increasingly powerful multimodal capabilities. However,

model-releasesarxiv-cs-cv
29 May 2026
Research

Do Language Models Track Entities Across State Changes?

DGX agent

arXiv:2605.30233v1 Announce Type: cross Abstract: Entity tracking (ET), the ability to keep track of states, is a fundamental skill that underlies complex reasoning. An increasing amount of work inves

researcharxiv-cs-ai
29 May 2026
Model Releases

Evaluating Cross-lingual Knowledge Consistency in Code-Mixed vis-a-vis Indian Languages using IndicKLAR

DGX agent

arXiv:2605.29637v1 Announce Type: new Abstract: Large language models recall knowledge reliably in English but often fail on the same query posed in a lower-resourced language -- a crosslingual consis

model-releasesarxiv-cs-cl
29 May 2026
Safety

EvoMD-LLM: Learning the Language of Species Evolution in Reactive Molecular Dynamics

DGX agent

arXiv:2605.29394v1 Announce Type: new Abstract: While large language models (LLMs) excel at static scientific reasoning, they struggle to model the temporal structure of dynamic physical processes. We

safetyarxiv-cs-ai
29 May 2026
Safety

From General Vision to Reliable Traversability Estimation: Adapting Vision Foundation Models for Unstructured Outdoor Environments

DGX agent

arXiv:2605.29565v1 Announce Type: new Abstract: Vision-based approaches have become the dominant paradigm for traversability estimation in unstructured outdoor environments, typically adapting vision

safetyarxiv-cs-cv
29 May 2026
Applications

GeoMag: Geometric-Aware Video Motion Magnification via State Space Model

DGX agent

arXiv:2605.29762v1 Announce Type: new Abstract: Video Motion Magnification (VMM) reveals imperceptible dynamics but often suffers from structural inconsistencies under complex geometric transformation

applicationsarxiv-cs-cv
29 May 2026
Safety

Geometry-Guided Modeling of Foundation Features Enables Generalizable Object Shape Deformation Learning

DGX agent

arXiv:2605.29661v1 Announce Type: new Abstract: Monocular 3D shape recovery is fundamental to geometric understanding, yet achieving robust generalization across arbitrary viewpoints and unseen object

safetyarxiv-cs-cv
29 May 2026
Research

GPS-Enhanced Tourist Mobility Modeling with Seasonal Spatial Priors and LLM-Based Activity Chain Generation

DGX agent

arXiv:2605.29578v1 Announce Type: new Abstract: Tourist mobility poses a distinct challenge for urban transportation planning. Unlike resident commuting, tourist travel is largely non-routine, attract

researcharxiv-cs-ai
29 May 2026
Model Releases

Hallucination Detection-Guided Preference Optimization for Clinical Summarization

DGX agent

arXiv:2605.28910v1 Announce Type: cross Abstract: Large language models (LLMs) have shown promise on summarization tasks, but they often produce hallucinations, which are unsupported or incorrect stat

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Hista and Numca: Estimate State Value Effectively for LLM Reinforcement Learning

DGX agent

arXiv:2605.29782v1 Announce Type: cross Abstract: Reinforcement learning (RL) refines large language models (LLMs) by directly optimizing model behavior through reward signals. While accurate state va

model-releasesarxiv-cs-ai
29 May 2026
Agents

Improving Collaborative Storytelling with a Multi-Agent Framework Based on Large Language Models

DGX agent

arXiv:2605.29625v1 Announce Type: new Abstract: The topic of Co-creation, i.e., AI agents interacting with humans to generate outputs (e.g., art), has gained significant attention recently. However, m

agentsarxiv-cs-ai
29 May 2026
Model Releases

Inform, Coach, Relate, Listen: Auditing LLM Caregiving Support Roles

DGX agent

arXiv:2605.29473v1 Announce Type: cross Abstract: Language models are increasingly being deployed for conversational support in informal caregiving contexts, where interactions often extend beyond inf

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Leak@k: Unlearning Does Not Make LLMs Forget Under Probabilistic Decoding

DGX agent

arXiv:2511.04934v3 Announce Type: replace Abstract: Unlearning in large language models (LLMs) is critical for regulatory compliance and for building ethical generative AI systems that avoid producing

model-releasesarxiv-cs-lg
29 May 2026
Research

Masked Diffusion Vision-Language Models for Temporal Action Localization

DGX agent

arXiv:2605.29858v1 Announce Type: new Abstract: Temporal action localization (TAL) requires recognizing the target event and localizing its start and end times precisely in untrimmed videos. Recent vi

researcharxiv-cs-cv
29 May 2026
Research

MMTM: Tri-Modal Topic Modeling for Long-Form Video via Similarity-Gated Fusion

DGX agent

arXiv:2605.29765v1 Announce Type: new Abstract: We introduce MMTM, a modular pipeline for topic discovery in long-form video that integrates speech recognition, audio and visual embeddings, and BERTop

researcharxiv-cs-lg
29 May 2026
Research

MVP-Shapley: Feature-based Modeling for Evaluating the Most Valuable Player in Basketball

DGX agent

arXiv:2506.04602v4 Announce Type: replace-cross Abstract: The burgeoning growth of the esports and multiplayer online gaming community has highlighted the critical importance of evaluating the Most Va

researcharxiv-cs-lg
29 May 2026
Safety

Permutation-Invariant Spectral Learning via Dyson Diffusion

DGX agent

arXiv:2510.08535v2 Announce Type: replace-cross Abstract: Diffusion models are central to generative modeling and have been adapted to graphs by diffusing adjacency matrix representations. The challen

safetyarxiv-cs-lg
29 May 2026
Model Releases

Realistic honeypot evaluations for scheming propensity

DGX agent

arXiv:2605.29729v1 Announce Type: new Abstract: We introduce scheming honeypot evaluations, a framework for testing whether models will pursue instrumental goals if given the opportunity. Our scheming

model-releasesarxiv-cs-lg
29 May 2026
Safety

ReasonLight: A Multimodal Foundation Model-Enhanced Reinforcement Learning Framework for Zero-Shot Traffic Signal Control

DGX agent

arXiv:2605.29425v1 Announce Type: new Abstract: Reinforcement learning (RL) has shown promise in traffic signal control (TSC). However, its reliance on predefined states limits responsiveness to obser

safetyarxiv-cs-ai
29 May 2026
Safety

SAHG: Sector-Anisotropic Hyperbolic Graph Model for Social Bot Detection

DGX agent

arXiv:2605.30166v1 Announce Type: cross Abstract: LLM-driven social bots can generate fluent, human-like text, reducing the discriminative advantage of content-based detection alone. However, coordina

safetyarxiv-cs-lg
29 May 2026
Safety

SigmaMedStat: Temporal Signal Modeling for ICU False Alarm Reduction

DGX agent

arXiv:2605.29236v1 Announce Type: new Abstract: Alarm fatigue in intensive care units (ICUs) is a well documented patient safety crisis. Clinical monitors generate 350 or more alarms per patient per d

safetyarxiv-cs-lg
29 May 2026
Model Releases

Small Agent Group is the Future of Digital Health

DGX agent

arXiv:2602.08013v2 Announce Type: replace Abstract: The rapid adoption of large language models (LLMs) in digital health has been driven by a 'scaling-first' philosophy, i.e., the assumption that clin

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

The Trust Paradox: How CS Researchers Engage LLM Leaderboards

DGX agent

arXiv:2605.28966v1 Announce Type: new Abstract: Large language model (LLM) leaderboards rank AI models using standardized benchmarks and have become highly visible across computer science, despite kno

model-releasesarxiv-cs-cl
29 May 2026
Research

Towards Continuous-time Causal Foundation Models

DGX agent

arXiv:2605.28880v1 Announce Type: new Abstract: Extending discrete-time causal Prior-data Fitted Networks for time series to continuous time invites writing the mechanism as a stochastic differential

researcharxiv-cs-lg
29 May 2026
Model Releases

AdaDPO: Self-Adaptive Direct Preference Optimization with Balanced Gradient Updates

DGX agent

arXiv:2605.28440v1 Announce Type: new Abstract: DPO has become a widely adopted alternative to RLHF for aligning LLMs with human preferences, eliminating the need for a separate reward model or RL loo

model-releasesarxiv-cs-cl
28 May 2026
Research

Applications of temporal graph learning for predicting the dynamics of biological systems

DGX agent

arXiv:2605.28659v1 Announce Type: new Abstract: Biological foundation models have shown strong performance in single-cell representation learning by applying transformer architectures directly to gene

researcharxiv-cs-lg
28 May 2026
Research

Debate Helps Weak Judges Reward Stronger Models

DGX agent

arXiv:2605.27483v1 Announce Type: cross Abstract: Despite theoretical promise, debate as a scalable oversight protocol has produced mixed empirical results: gains in some settings, and null effects in

researcharxiv-cs-ai
28 May 2026
Safety

DebFilter: Eradicating Biases Stashed in Value

DGX agent

arXiv:2605.28167v1 Announce Type: new Abstract: Text-to-image diffusion models, which are theoretically equivalent to score-based generative models, generate images through a multi-step denoising proc

safetyarxiv-cs-cv
28 May 2026
Model Releases

FPMoE: A Sparse Mixture-of-Experts Approach to Functional Code Generation

DGX agent

arXiv:2605.27849v1 Announce Type: cross Abstract: Despite rapid progress in LLM-based code generation, existing models are predominantly trained on imperative languages, leaving functional programming

model-releasesarxiv-cs-ai
28 May 2026
Agents

Heterogeneous Multi-Agent Modeling for Measurement and Network Analysis of the Data Service Market

DGX agent

arXiv:2605.27433v1 Announce Type: cross Abstract: With the increasing complexity of collaboration among various social entities and user demands, the factors affecting the stable development of the da

agentsarxiv-cs-ai
28 May 2026
Research

Hybrid Neural World Models

DGX agent

arXiv:2605.28317v1 Announce Type: cross Abstract: Neural surrogates promise large speedups over classical solvers for physical dynamics but fail silently at sharp dynamical events such as shocks, fron

researcharxiv-cs-ai
28 May 2026
Model Releases

KVoiceBench, KOpenAudioBench, and KMMAU: Agent-Driven Korean Speech Benchmarks for Evaluating SpeechLMs

DGX agent

arXiv:2605.27984v1 Announce Type: cross Abstract: Speech language models (SpeechLMs) have achieved substantial progress by extending large language models (LLMs) to the speech modality. However, Speec

model-releasesarxiv-cs-ai
28 May 2026
Tutorials

Learning the Error Patterns of Language Models

DGX agent

arXiv:2605.28328v1 Announce Type: cross Abstract: When generating outputs for domains with specific validity constraints (e.g., a program should compile), LLMs often fail in a small number of focused

tutorialsarxiv-cs-ai
28 May 2026
Safety

Mathematical Modelling of Ethical AI Use in Higher Education: A Coordination Game Framework for Future-Facing Learning

DGX agent

arXiv:2605.27400v1 Announce Type: cross Abstract: The rapid uptake of generative artificial intelligence (AI) in higher education is reshaping assessment practices and intensifying concerns around aca

safetyarxiv-cs-ai
28 May 2026
← Previous
1…261262263264265…1058
Next →