AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,457
  • Agents7,399
  • Applications5,302
  • Concepts5
  • Hardware1,786
  • Industry6,117
  • Local Ai4,835
  • Model Releases23,193
  • Research19,715
  • Safety13,094
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,457
  • Agents7,399
  • Applications5,302
  • Concepts5
  • Hardware1,786
  • Industry6,117
  • Local Ai4,835
  • Model Releases23,193
  • Research19,715
  • Safety13,094
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
86,457Total entries
1Added by human
86,456Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
50,764 results
Model Releases

See Further, Think Deeper: Advancing VLM's Reasoning Ability with Low-level Visual Cues and Reflection

DGX agent

arXiv:2604.24339v1 Announce Type: cross Abstract: Recent advances in Vision-Language Models (VLMs) have benefited from Reinforcement Learning (RL) for enhanced reasoning. However, existing methods sti

model-releasesarxiv-cs-ai
28 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

SycoPhantasy: Quantifying Sycophancy and Hallucination in Small Open Weight VLMs for Vision-Language Scoring of Fantasy Characters

DGX agent

arXiv:2604.24346v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly deployed as evaluators in tasks requiring nuanced image understanding, yet their reliability in scoring

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

TextGround4M: A Prompt-Aligned Dataset for Layout-Aware Text Rendering

DGX agent

arXiv:2604.24459v1 Announce Type: new Abstract: Despite recent advances in text-to-image generation, models still struggle to accurately render prompt-specified text with correct spatial layout -- esp

model-releasesarxiv-cs-cv
28 Apr 2026
Research

Uncertainty Propagation in LLM-Based Systems

DGX agent

arXiv:2604.23505v1 Announce Type: cross Abstract: Uncertainty in large language model (LLM)-based systems is often studied at the level of a single model output, yet deployed LLM applications are comp

researcharxiv-cs-ai
28 Apr 2026
Safety

Vision-Language-Action Safety: Threats, Challenges, Evaluations, and Mechanisms

DGX agent

arXiv:2604.23775v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models are emerging as a unified substrate for embodied intelligence. This shift raises a new class of safety challenges, s

safetyarxiv-cs-ro
28 Apr 2026
Research

Multi-Token Prediction via Self-Distillation

DGX agent

arXiv:2602.06019v2 Announce Type: replace Abstract: Existing techniques for accelerating language model inference, such as speculative decoding, require training auxiliary speculator models and buildi

researcharxiv-cs-cl
27 Apr 2026
Safety

PermaFrost-Attack: Stealth Pretraining Seeding(SPS) for planting Logic Landmines During LLM Training

DGX agent

arXiv:2604.22117v1 Announce Type: cross Abstract: Aligned large language models(LLMs) remain vulnerable to adversarial manipulation, and their dependence on web-scale pretraining creates a subtle but

safetyarxiv-cs-ai
27 Apr 2026
Research

Privacy Leakage via Output Label Space and Differentially Private Continual Learning

DGX agent

arXiv:2411.04680v5 Announce Type: replace Abstract: Differential privacy (DP) is a formal privacy framework that enables training machine learning (ML) models while protecting individuals' data. As po

researcharxiv-cs-lg
27 Apr 2026
Model Releases

SHAPE: Unifying Safety, Helpfulness and Pedagogy for Educational LLMs

DGX agent

arXiv:2604.22134v1 Announce Type: new Abstract: Large Language Models (LLMs) have been widely explored in educational scenarios. We identify a critical vulnerability in current educational LLMs, pedag

model-releasesarxiv-cs-cl
27 Apr 2026
Model Releases

TS-Arena -- A Live Forecast Pre-Registration Platform

DGX agent

arXiv:2512.20761v3 Announce Type: replace-cross Abstract: Time Series Foundation Models (TSFMs) are transforming the field of forecasting. However, evaluating them on historical data is increasingly d

model-releasesarxiv-cs-ai
27 Apr 2026
Research

Using Embedding Models to Improve Probabilistic Race Prediction

DGX agent

arXiv:2604.22555v1 Announce Type: new Abstract: Estimating racial disparity requires individual-level race data, which are often unavailable due to the sensitivity of collecting such information. To a

researcharxiv-cs-cl
27 Apr 2026
Model Releases

ADS-POI: Agentic Spatiotemporal State Decomposition for Next Point-of-Interest Recommendation

DGX agent

arXiv:2604.20846v1 Announce Type: cross Abstract: Next point-of-interest (POI) recommendation requires modeling user mobility as a spatiotemporal sequence, where different behavioral factors may evolv

model-releasesarxiv-cs-ai
24 Apr 2026
Research

Analytical FFN-to-MoE Restructuring via Activation Pattern Analysis

DGX agent

arXiv:2502.04416v3 Announce Type: replace-cross Abstract: Scaling large language models (LLMs) improves performance but significantly increases inference costs, with feed-forward networks (FFNs) consu

researcharxiv-cs-ai
24 Apr 2026
Model Releases

Breaking Bad: Interpretability-Based Safety Audits of State-of-the-Art LLMs

DGX agent

arXiv:2604.20945v1 Announce Type: cross Abstract: Effective safety auditing of large language models (LLMs) demands tools that go beyond black-box probing and systematically uncover vulnerabilities ro

model-releasesarxiv-cs-lg
24 Apr 2026
Research

Clinically-Informed Modeling for Pediatric Brain Tumor Classification from Whole-Slide Histopathology Images

DGX agent

arXiv:2604.21060v1 Announce Type: new Abstract: Accurate diagnosis of pediatric brain tumors, starting with histopathology, presents unique challenges for deep learning, including severe data scarcity

researcharxiv-cs-cv
24 Apr 2026
Model Releases

GeoRA: Geometry-Aware Low-Rank Adaptation for RLVR

DGX agent

arXiv:2601.09361v3 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) is a key paradigm for improving large-scale reasoning models. Unlike supervised fine-tun

model-releasesarxiv-cs-ai
24 Apr 2026
Safety

How VLAs (Really) Work In Open-World Environments

DGX agent

arXiv:2604.21192v1 Announce Type: cross Abstract: Vision-language-action models (VLAs) have been extensively used in robotics applications, achieving great success in various manipulation problems. Mo

safetyarxiv-cs-ai
24 Apr 2026
Model Releases

Hyperloop Transformers

DGX agent

arXiv:2604.21254v1 Announce Type: cross Abstract: LLM architecture research generally aims to maximize model quality subject to fixed compute/latency budgets. However, many applications of interest su

model-releasesarxiv-cs-cl
24 Apr 2026
Research

Not-a-Bandit: Provably No-Regret Drafter Selection in Speculative Decoding for LLMs

DGX agent

arXiv:2510.20064v2 Announce Type: replace Abstract: Speculative decoding is widely used in accelerating large language model (LLM) inference. In this work, we focus on the online draft model selection

researcharxiv-cs-lg
24 Apr 2026
Model Releases

OpenEstimate: Evaluating LLMs on Reasoning Under Uncertainty with Real-World Data

DGX agent

arXiv:2510.15096v2 Announce Type: replace Abstract: Real-world settings where language models (LMs) are deployed -- in domains spanning healthcare, finance, and other forms of knowledge work -- requir

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Separable Expert Architecture: Toward Privacy-Preserving LLM Personalization via Composable Adapters and Deletable User Proxies

DGX agent

arXiv:2604.21571v1 Announce Type: new Abstract: Current model training approaches incorporate user information directly into shared weights, making individual data removal computationally infeasible w

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Towards Universal Tabular Embeddings: A Benchmark Across Data Tasks

DGX agent

arXiv:2604.21696v1 Announce Type: new Abstract: Tabular foundation models aim to learn universal representations of tabular data that transfer across tasks and domains, enabling applications such as t

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

VARestorer: One-Step VAR Distillation for Real-World Image Super-Resolution

DGX agent

arXiv:2604.21450v1 Announce Type: cross Abstract: Recent advancements in visual autoregressive models (VAR) have demonstrated their effectiveness in image generation, highlighting their potential for

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

VistaBot: View-Robust Robot Manipulation via Spatiotemporal-Aware View Synthesis

DGX agent

arXiv:2604.21914v1 Announce Type: new Abstract: Recently, end-to-end robotic manipulation models have gained significant attention for their generalizability and scalability. However, they often suffe

model-releasesarxiv-cs-ro
24 Apr 2026
Model Releases

Wiring the 'Why': A Unified Taxonomy and Survey of Abductive Reasoning in LLMs

DGX agent

arXiv:2604.08016v2 Announce Type: replace Abstract: Regardless of its foundational role in human discovery and sense-making, abductive reasoning--the inference of the most plausible explanation for an

model-releasesarxiv-cs-ai
24 Apr 2026
Research

AFMRL: Attribute-Enhanced Fine-Grained Multi-Modal Representation Learning in E-commerce

DGX agent

arXiv:2604.20135v1 Announce Type: new Abstract: Multimodal representation is crucial for E-commerce tasks such as identical product retrieval. Large representation models (e.g., VLM2Vec) demonstrate s

researcharxiv-cs-cl
23 Apr 2026
Model Releases

CCTVBench: Contrastive Consistency Traffic VideoQA Benchmark for Multimodal LLMs

DGX agent

arXiv:2604.20460v1 Announce Type: new Abstract: Safety-critical traffic reasoning requires contrastive consistency: models must detect true hazards when an accident occurs, and reliably reject plausib

model-releasesarxiv-cs-cv
23 Apr 2026
Safety

DAIRE: A lightweight AI model for real-time detection of Controller Area Network attacks in the Internet of Vehicles

DGX agent

arXiv:2604.20771v1 Announce Type: cross Abstract: The Internet of Vehicles (IoV) is advancing modern transportation by improving safety, efficiency, and intelligence. However, the reliance on the Cont

safetyarxiv-cs-ai
23 Apr 2026
Applications

Do Hallucination Neurons Generalize? Evidence from Cross-Domain Transfer in LLMs

DGX agent

arXiv:2604.19765v1 Announce Type: cross Abstract: Recent work identifies a sparse set of 'hallucination neurons' (H-neurons), less than 0.1% of feed-forward network neurons, that reliably predict when

applicationsarxiv-cs-ai
23 Apr 2026
Model Releases

Evian: Towards Explainable Visual Instruction-tuning Data Auditing

DGX agent

arXiv:2604.20544v1 Announce Type: cross Abstract: The efficacy of Large Vision-Language Models (LVLMs) is critically dependent on the quality of their training data, requiring a precise balance betwee

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Exploring Spatial Intelligence from a Generative Perspective

DGX agent

arXiv:2604.20570v1 Announce Type: new Abstract: Spatial intelligence is essential for multimodal large language models, yet current benchmarks largely assess it only from an understanding perspective.

model-releasesarxiv-cs-cv
23 Apr 2026
Safety

Generative Augmentation of Imbalanced Flight Records for Flight Diversion Prediction: A Multi-objective Optimisation Framework

DGX agent

arXiv:2604.20288v1 Announce Type: new Abstract: Flight diversions are rare but high-impact events in aviation, making their reliable prediction vital for both safety and operational efficiency. Howeve

safetyarxiv-cs-lg
23 Apr 2026
Model Releases

LoRA-FA: Efficient and Effective Low Rank Representation Fine-tuning

DGX agent

arXiv:2308.03303v2 Announce Type: replace Abstract: Fine-tuning large language models (LLMs) is crucial for improving their performance on downstream tasks, but full-parameter fine-tuning (Full-FT) is

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

Self-Awareness before Action: Mitigating Logical Inertia via Proactive Cognitive Awareness

DGX agent

arXiv:2604.20413v1 Announce Type: new Abstract: Large language models perform well on many reasoning tasks, yet they often lack awareness of whether their current knowledge or reasoning state is compl

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Understanding the Staged Dynamics of Transformers in Learning Latent Structure

DGX agent

arXiv:2511.19328v2 Announce Type: replace Abstract: Language modeling has shown us that transformers can discover latent structure from context, but the dynamics of how they acquire different componen

model-releasesarxiv-cs-lg
23 Apr 2026
Applications

Wan-Image: Pushing the Boundaries of Generative Visual Intelligence

DGX agent

arXiv:2604.19858v1 Announce Type: new Abstract: We present Wan-Image, a unified visual generation system explicitly engineered to paradigm-shift image generation models from casual synthesizers into p

applicationsarxiv-cs-cv
23 Apr 2026
Model Releases

X-PCR: A Benchmark for Cross-modality Progressive Clinical Reasoning in Ophthalmic Diagnosis

DGX agent

arXiv:2604.20350v1 Announce Type: new Abstract: Despite significant progress in Multi-modal Large Language Models (MLLMs), their clinical reasoning capacity for multi-modal diagnosis remains largely u

model-releasesarxiv-cs-cv
23 Apr 2026
Safety

Beyond Semantic Similarity: A Component-Wise Evaluation Framework for Medical Question Answering Systems with Health Equity Implications

DGX agent

arXiv:2604.19281v1 Announce Type: cross Abstract: The use of Large Language Models (LLMs) to support patients in addressing medical questions is becoming increasingly prevalent. However, most of the m

safetyarxiv-cs-ai
22 Apr 2026
Model Releases

Compile to Compress: Boosting Formal Theorem Provers by Compiler Outputs

DGX agent

arXiv:2604.18587v1 Announce Type: cross Abstract: Large language models (LLMs) have demonstrated significant potential in formal theorem proving, yet state-of-the-art performance often necessitates pr

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Council Mode: Mitigating Hallucination and Bias in LLMs via Multi-Agent Consensus

DGX agent

arXiv:2604.02923v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs), particularly those employing Mixture-of-Experts (MoE) architectures, have achieved remarkable capabilities acros

model-releasesarxiv-cs-ai
22 Apr 2026
Local Ai

Distillation Traps and Guards: A Calibration Knob for LLM Distillability

DGX agent

arXiv:2604.18963v1 Announce Type: cross Abstract: Knowledge distillation (KD) transfers capabilities from large language models (LLMs) to smaller students, yet it can fail unpredictably and also under

local-aiarxiv-cs-ai
22 Apr 2026
Safety

HALO: Hybrid Auto-encoded Locomotion with Learned Latent Dynamics, Poincare Maps, and Regions of Attraction

DGX agent

arXiv:2604.18887v1 Announce Type: new Abstract: Reduced-order models are powerful for analyzing and controlling high-dimensional dynamical systems. Yet constructing these models for complex hybrid sys

safetyarxiv-cs-ro
22 Apr 2026
Applications

IMPACT: Importance-Aware Activation Space Reconstruction

DGX agent

arXiv:2507.03828v4 Announce Type: replace Abstract: Large language models (LLMs) achieve strong performance across diverse domains but remain difficult to deploy in resource-constrained environments d

applicationsarxiv-cs-lg
22 Apr 2026
Research

Model-Agnostic Meta Learning for Class Imbalance Adaptation

DGX agent

arXiv:2604.18759v1 Announce Type: new Abstract: Class imbalance is a widespread challenge in NLP tasks, significantly hindering robust performance across diverse domains and applications. We introduce

researcharxiv-cs-cl
22 Apr 2026
Safety

Multi-Task Reinforcement Learning for Enhanced Multimodal LLM-as-a-Judge

DGX agent

arXiv:2603.11665v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have been widely adopted as MLLM-as-a-Judges due to their strong alignment with human judgment across vario

safetyarxiv-cs-cl
22 Apr 2026
Research

Real-Time Streamable Generative Speech Restoration with Flow Matching

DGX agent

arXiv:2512.19442v3 Announce Type: replace-cross Abstract: Diffusion-based generative models have greatly impacted the speech processing field in recent years, exhibiting high speech naturalness and sp

researcharxiv-cs-lg
22 Apr 2026
Model Releases

SitEmb-v1.5: Improved Context-Aware Dense Retrieval for Semantic Association and Long Story Comprehension

DGX agent

arXiv:2508.01959v2 Announce Type: replace Abstract: Retrieval-augmented generation (RAG) over long documents typically involves splitting the text into smaller chunks, which serve as the basic units f

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

Towards Understanding the Robustness of Sparse Autoencoders

DGX agent

arXiv:2604.18756v1 Announce Type: cross Abstract: Large Language Models (LLMs) remain vulnerable to optimization-based jailbreak attacks that exploit internal gradient structure. While Sparse Autoenco

model-releasesarxiv-cs-ai
22 Apr 2026
← Previous
1…269270271272273…1058
Next →