AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,435 results
Model Releases

When to Retrieve During Reasoning: Adaptive Retrieval for Large Reasoning Models

DGX agent

arXiv:2604.26649v1 Announce Type: cross Abstract: Large reasoning models such as DeepSeek-R1 and OpenAI o1 generate extended chains of thought spanning thousands of tokens, yet their integration with

model-releasesarxiv-cs-ai
30 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

EvoTSC: Evolving Feature Learning Models for Time Series Classification via Genetic Programming

DGX agent

arXiv:2604.25499v1 Announce Type: new Abstract: Time series classification is an important analytical task across diverse domains. However, its practical application is often hindered by the scarcity

model-releasesarxiv-cs-lg
29 Apr 2026
Model Releases

Exploring Reasoning Reward Model for Agents

DGX agent

arXiv:2601.22154v2 Announce Type: replace-cross Abstract: Agentic Reinforcement Learning (Agentic RL) has achieved notable success in enabling agents to perform complex reasoning and tool use. However

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

Faithfulness-QA: A Counterfactual Entity Substitution Dataset for Training Context-Faithful RAG Models

DGX agent

arXiv:2604.25313v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) models frequently produce answers grounded in parametric memory rather than the retrieved context, undermining the

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

From Local to Global: Revisiting Structured Pruning Paradigms for Large Language Models

DGX agent

arXiv:2510.18030v2 Announce Type: replace Abstract: Structured pruning is a practical approach to deploying large language models (LLMs) efficiently, as it yields compact, hardware-friendly architectu

model-releasesarxiv-cs-cl
29 Apr 2026
Agents

From Soliloquy to Agora: Memory-Enhanced LLM Agents with Decentralized Debate for Optimization Modeling

DGX agent

arXiv:2604.25847v1 Announce Type: cross Abstract: Optimization modeling underpins real-world decision-making in logistics, manufacturing, energy, and public services, but reliably solving such problem

agentsarxiv-cs-lg
29 Apr 2026
Applications

MotionBricks: Scalable Real-Time Motions with Modular Latent Generative Model and Smart Primitives

DGX agent

arXiv:2604.24833v1 Announce Type: cross Abstract: Despite transformative advances in generative motion synthesis, real-time interactive motion control remains dominated by traditional techniques. In t

applicationsarxiv-cs-lg
29 Apr 2026
Model Releases

Phase-Associative Memory: Sequence Modeling in Complex Hilbert Space

DGX agent

arXiv:2604.05030v2 Announce Type: replace Abstract: Experiments probing natural language processing by both humans and LLMs suggest that the meaning of a semantic expression is indeterminate prior to

model-releasesarxiv-cs-cl
29 Apr 2026
Safety

Three Models of RLHF Annotation: Extension, Evidence, and Authority

DGX agent

arXiv:2604.25895v1 Announce Type: cross Abstract: Preference-based alignment methods, most prominently Reinforcement Learning with Human Feedback (RLHF), use the judgments of human annotators to shape

safetyarxiv-cs-cl
29 Apr 2026
Model Releases

A Limit Theory of Foundation Models: A Mathematical Approach to Understanding Emergent Intelligence and Scaling Laws

DGX agent

arXiv:2604.24037v1 Announce Type: new Abstract: Emergent intelligence have played a major role in the modern AI development. While existing studies primarily rely on empirical observations to characte

model-releasesarxiv-cs-lg
28 Apr 2026
Research

A Survey on Split Learning for LLM Fine-Tuning: Models, Systems, and Privacy Optimizations

DGX agent

arXiv:2604.24468v1 Announce Type: cross Abstract: Fine-tuning unlocks large language models (LLMs) for specialized applications, but its high computational cost often puts it out of reach for resource

researcharxiv-cs-cl
28 Apr 2026
Research

Accelerating Frequency Domain Diffusion Models with Error-Feedback Event-Driven Caching

DGX agent

arXiv:2604.22901v1 Announce Type: new Abstract: Diffusion models achieve remarkable success in time series generation. However, slow inference limits their practical deployment. We propose E^2-CRF (Er

researcharxiv-cs-lg
28 Apr 2026
Model Releases

AIPsy-Affect: A Keyword-Free Clinical Stimulus Battery for Mechanistic Interpretability of Emotion in Language Models

DGX agent

arXiv:2604.23719v1 Announce Type: cross Abstract: Mechanistic interpretability research on emotion in large language models -- linear probing, activation patching, sparse autoencoder (SAE) feature ana

model-releasesarxiv-cs-ai
28 Apr 2026
Safety

An Information-Geometric Framework for Stability Analysis of Large Language Models under Entropic Stress

DGX agent

arXiv:2604.24076v1 Announce Type: new Abstract: As large language models (LLMs) are increasingly deployed in high-stakes and operational settings, evaluation strategies based solely on aggregate accur

safetyarxiv-cs-ai
28 Apr 2026
Model Releases

AutoPyVerifier: Learning Compact Executable Verifiers for Large Language Model Outputs

DGX agent

arXiv:2604.22937v1 Announce Type: new Abstract: Verification is becoming central to both reinforcement-learning-based training and inference-time control of large language models (LLMs). Yet current v

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

Benchmarking and Mitigating Sycophancy in Medical Vision Language Models

DGX agent

arXiv:2509.21979v4 Announce Type: replace-cross Abstract: Visual language models (VLMs) have the potential to transform medical workflows. However, the deployment is limited by sycophancy. Despite thi

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Can Multimodal Large Language Models Truly Understand Small Objects?

DGX agent

arXiv:2604.22884v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have shown promising potential in diverse understanding tasks, e.g., image and video analysis, math and physi

model-releasesarxiv-cs-ai
28 Apr 2026
Research

Characterizing Vision-Language-Action Models across XPUs: Constraints and Acceleration for On-Robot Deployment

DGX agent

arXiv:2604.24447v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models are promising for generalist robot control, but on-robot deployment is bottlenecked by real-time inference under t

researcharxiv-cs-ai
28 Apr 2026
Research

Contextual Linear Activation Steering of Language Models

DGX agent

arXiv:2604.24693v1 Announce Type: new Abstract: Linear activation steering is a powerful approach for eliciting the capabilities of large language models and specializing their behavior using limited

researcharxiv-cs-cl
28 Apr 2026
Safety

Discovering Failure Modes in Vision-Language Models using RL

DGX agent

arXiv:2604.04733v2 Announce Type: replace-cross Abstract: Vision-language Models (VLMs), despite achieving strong performance on multimodal benchmarks, often misinterpret straightforward visual concep

safetyarxiv-cs-ai
28 Apr 2026
Model Releases

DO-Bench: An Attributable Benchmark for Diagnosing Object Hallucination in Vision-Language Models

DGX agent

arXiv:2604.22822v1 Announce Type: cross Abstract: Object level hallucination remains a central reliability challenge for vision language models (VLMs), particularly in binary object existence verifica

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

EAGLE: Expert-Augmented Attention Guidance for Tuning-Free Industrial Anomaly Detection in Multimodal Large Language Models

DGX agent

arXiv:2602.17419v3 Announce Type: replace Abstract: Multimodal large language models (MLLMs) can enrich industrial anomaly detection with semantic descriptions and anomaly reasoning, but they still la

model-releasesarxiv-cs-cv
28 Apr 2026
Applications

FedRef: Bayesian Fine-Tuning using a Reference Model to Mitigate Catastrophic Forgetting for Heterogeneous Federated Learning

DGX agent

arXiv:2506.23210v5 Announce Type: replace-cross Abstract: Federated learning (FL) enables collaborative model training across distributed clients while preserving data privacy. However, data and syste

applicationsarxiv-cs-ai
28 Apr 2026
Tutorials

Fine-tuning vs. In-context Learning in Large Language Models: A Formal Language Learning Perspective

DGX agent

arXiv:2604.23267v1 Announce Type: new Abstract: Large language models (LLMs) operate in two fundamental learning modes - fine-tuning (FT) and in-context learning (ICL) - raising key questions about wh

tutorialsarxiv-cs-cl
28 Apr 2026
Research

HeadRouter: Dynamic Head-Weight Routing for Task-Adaptive Audio Token Pruning in Large Audio Language Models

DGX agent

arXiv:2604.23717v1 Announce Type: cross Abstract: Recent large audio language models (LALMs) demonstrate remarkable capabilities in processing extended multi-modal sequences, yet incur high inference

researcharxiv-cs-cl
28 Apr 2026
Applications

HeiSD: Hybrid Speculative Decoding for Embodied Vision-Language-Action Models with Kinematic Awareness

DGX agent

arXiv:2603.17573v2 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) Models have become the mainstream solution for robot control, but suffer from slow inference speeds. Speculative

applicationsarxiv-cs-lg
28 Apr 2026
Research

Instruction-Free Tuning of Large Vision Language Models for Medical Instruction Following

DGX agent

arXiv:2603.19482v2 Announce Type: replace Abstract: Large vision language models (LVLMs) have demonstrated impressive performance across a wide range of tasks. These capabilities largely stem from vis

researcharxiv-cs-cv
28 Apr 2026
Tutorials

Inverting Foundation Models of Brain Function with Simulation-Based Inference

DGX agent

arXiv:2604.23865v1 Announce Type: cross Abstract: Foundation models of brain activity promise a new frontier for in silico neuroscience by emulating neural responses to complex stimuli across tasks an

tutorialsarxiv-cs-ai
28 Apr 2026
Model Releases

Large Language Models as Virtual Survey Respondents: Evaluating Sociodemographic Response Generation

DGX agent

arXiv:2509.06337v2 Announce Type: replace Abstract: Questionnaire-based surveys are foundational to social science research and public policymaking, yet traditional survey methods remain costly, time-

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Less Is More: Engineering Challenges of On-Device Small Language Model Integration in a Mobile Application

DGX agent

arXiv:2604.24636v1 Announce Type: cross Abstract: On-device Small Language Models (SLMs) promise fully offline, private AI experiences for mobile users (no cloud dependency, no data leaving the device

model-releasesarxiv-cs-ai
28 Apr 2026
Research

Live Knowledge Tracing: Real-Time Adaptation using Tabular Foundation Models

DGX agent

arXiv:2602.06542v3 Announce Type: replace Abstract: Deep knowledge tracing models have achieved significant breakthroughs in modeling student learning trajectories. However, these architectures requir

researcharxiv-cs-lg
28 Apr 2026
Model Releases

Machine Learning and Deep Learning Models for Short Term Electricity Price Forecasting in Australia's National Electricity Market

DGX agent

arXiv:2604.23908v1 Announce Type: new Abstract: Short term electricity price forecast is essential in competitive power markets, yet electricity price series exhibit high volatility, irregularity, and

model-releasesarxiv-cs-lg
28 Apr 2026
Research

Machine learning models for estimating counterfactuals in a single-arm inflammatory bowel disease study

DGX agent

arXiv:2604.23465v1 Announce Type: new Abstract: Single-arm trials accelerate study timelines by reducing the number of patients that must be recruited for a concurrent control group. However, these de

researcharxiv-cs-lg
28 Apr 2026
Research

NVILA: Efficient Frontier Visual Language Models

DGX agent

arXiv:2412.04468v3 Announce Type: replace Abstract: Visual language models (VLMs) have made significant advances in accuracy in recent years. However, their efficiency has received much less attention

researcharxiv-cs-cv
28 Apr 2026
Model Releases

PDF-WuKong: A Large Multimodal Model for Efficient Long PDF Reading with End-to-End Sparse Sampling

DGX agent

arXiv:2410.05970v3 Announce Type: replace-cross Abstract: Multimodal document understanding is a challenging task to process and comprehend large amounts of textual and visual information. Recent adva

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Seeing Is No Longer Believing: Frontier Image Generation Models, Synthetic Visual Evidence, and Real-World Risk

DGX agent

arXiv:2604.24197v1 Announce Type: cross Abstract: Frontier image generation has moved from artistic synthesis toward synthetic visual evidence. Systems such as GPT Image 2, Nano Banana Pro, Nano Banan

model-releasesarxiv-cs-ai
28 Apr 2026
Research

Small Language Model Helps Resolve Semantic Ambiguity of LLM Prompt

DGX agent

arXiv:2604.23263v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly utilized in various complex reasoning tasks due to their excellent instruction following capability. How

researcharxiv-cs-ai
28 Apr 2026
Model Releases

Speech Enhancement Based on Drifting Models

DGX agent

arXiv:2604.24199v1 Announce Type: cross Abstract: We propose Speech Enhancement based on Drifting Models (DriftSE), a novel generative framework that formulates denoising as an equilibrium problem. Ra

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Think-at-Hard: Selective Latent Iterations to Improve Reasoning Language Models

DGX agent

arXiv:2511.08577v2 Announce Type: replace-cross Abstract: Improving reasoning abilities of Large Language Models (LLMs), especially under parameter constraints, is crucial for real-world applications.

model-releasesarxiv-cs-ai
28 Apr 2026
Local Ai

Unified Multi-Foundation-Model Slide Representation for Pan-Cancer Recognition and Text-Guided Tumor Localization

DGX agent

arXiv:2604.22846v1 Announce Type: new Abstract: The expanding ecosystem of pathology foundation models has produced powerful but fragmented tile-level representations, limiting their use in clinical t

local-aiarxiv-cs-cv
28 Apr 2026
Model Releases

Can Large Language Models Adequately Perform Symbolic Reasoning Over Time Series?

DGX agent

arXiv:2508.03963v4 Announce Type: replace Abstract: Uncovering hidden symbolic laws from time series data, as an aspiration dating back to Kepler's discovery of planetary motion, remains a core challe

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

FILTR: Extracting Topological Features from Pretrained 3D Models

DGX agent

arXiv:2604.22334v1 Announce Type: new Abstract: Recent advances in pretraining 3D point cloud encoders (e.g., Point-BERT, Point-MAE) have produced powerful models, whose abilities are typically evalua

model-releasesarxiv-cs-cv
27 Apr 2026
Local Ai

Fine-Grained Analysis of Shared Syntactic Mechanisms in Language Models

DGX agent

arXiv:2604.22166v1 Announce Type: new Abstract: While language models demonstrate sophisticated syntactic capabilities, the extent to which their internal mechanisms align with cross-constructional pr

local-aiarxiv-cs-cl
27 Apr 2026
Model Releases

From Interpretability to Performance: Optimizing Retrieval Heads for Long-Context Language Models

DGX agent

arXiv:2601.11020v3 Announce Type: replace Abstract: Advances in mechanistic interpretability have identified special attention heads, known as retrieval heads, that are responsible for retrieving info

model-releasesarxiv-cs-cl
27 Apr 2026
Model Releases

Graph-to-Vision: Multi-graph Understanding and Reasoning using Vision-Language Models

DGX agent

arXiv:2503.21435v3 Announce Type: replace Abstract: Recent advances in Vision-Language Models (VLMs) have shown promising capabilities in interpreting visualized graph data, offering a new perspective

model-releasesarxiv-cs-ai
27 Apr 2026
Local Ai

MambaCSP: Hybrid-Attention State Space Models for Hardware-Efficient Channel State Prediction

DGX agent

arXiv:2604.21957v1 Announce Type: cross Abstract: Recent works have demonstrated that attention-based transformer and large language model (LLM) architectures can achieve strong channel state predicti

local-aiarxiv-cs-ai
27 Apr 2026
Applications

Mochi: Aligning Pre-training and Inference for Efficient Graph Foundation Models via Meta-Learning

DGX agent

arXiv:2604.22031v1 Announce Type: cross Abstract: We propose Mochi, a Graph Foundation Model that addresses task unification and training efficiency by adopting a meta-learning based training framewor

applicationsarxiv-cs-ai
27 Apr 2026
Model Releases

MTT-Bench: Predicting Social Dominance in Mice via Multimodal Large Language Models

DGX agent

arXiv:2604.22492v1 Announce Type: cross Abstract: Understanding social dominance in animal behavior is critical for neuroscience and behavioral studies. In this work, we explore the capability of Mult

model-releasesarxiv-cs-cv
27 Apr 2026
← Previous
1…9192939495…1030
Next →