AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,435 results
Research

Reducing Peak Memory Usage for Modern Multimodal Large Language Model Pipelines

DGX agent

arXiv:2604.16734v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have recently demonstrated strong capabilities in understanding and generating responses from diverse visual in

researcharxiv-cs-cv
21 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Applications

REFLEX: Reference-Free Evaluation of Log Summarization via Large Language Model Judgment

DGX agent

arXiv:2511.07458v2 Announce Type: replace Abstract: Evaluating log summarization systems is challenging due to the lack of high-quality reference summaries and the limitations of existing metrics like

applicationsarxiv-cs-cl
21 Apr 2026
Model Releases

Representation Before Training: A Fixed-Budget Benchmark for Generative Medical Event Models

DGX agent

arXiv:2604.16775v1 Announce Type: new Abstract: Every prediction from a generative medical event model is bounded by how clinical events are tokenized, yet input representation is rarely isolated from

model-releasesarxiv-cs-lg
21 Apr 2026
Safety

Retrieval-Augmented Multimodal Model for Fake News Detection

DGX agent

arXiv:2604.18112v1 Announce Type: new Abstract: In recent years, multimodal multidomain fake news detection has garnered increasing attention. Nevertheless, this direction presents two significant cha

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

Revisiting a Pain in the Neck: A Semantic Reasoning Benchmark for Language Models

DGX agent

arXiv:2604.16593v1 Announce Type: new Abstract: We present SemanticQA, an evaluation suite designed to assess language models (LMs) in semantic phrase processing tasks. The benchmark consolidates exis

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

SHRUG-FM: Reliability-Aware Foundation Models for Earth Observation

DGX agent

arXiv:2511.10370v2 Announce Type: replace Abstract: Geospatial foundation models (GFMs) for Earth observation often fail to perform reliably in environments underrepresented during pretraining. We int

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

SpeechMedAssist: Efficiently and Effectively Adapting Speech Language Models for Medical Consultation

DGX agent

arXiv:2601.04638v2 Announce Type: replace Abstract: Medical consultations are intrinsically speech-centric. However, most prior works focus on long-text-based interactions, which are cumbersome and pa

model-releasesarxiv-cs-cl
21 Apr 2026
Safety

SPS: Steering Probability Squeezing for Better Exploration in Reinforcement Learning for Large Language Models

DGX agent

arXiv:2604.16995v1 Announce Type: new Abstract: Reinforcement learning (RL) has emerged as a promising paradigm for training reasoning-oriented models by leveraging rule-based reward signals. However,

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement

DGX agent

arXiv:2604.17887v1 Announce Type: new Abstract: Inverse Dynamics Models (IDMs) map visual observations to low-level action commands, serving as central components for data labeling and policy executio

model-releasesarxiv-cs-ro
21 Apr 2026
Research

StableMTL: Repurposing Latent Diffusion Models for Multi-Task Learning from Partially Annotated Synthetic Datasets

DGX agent

arXiv:2506.08013v2 Announce Type: replace Abstract: Multi-task learning for dense prediction is limited by the need for extensive annotation for every task, though recent works have explored training

researcharxiv-cs-cv
21 Apr 2026
Research

StrEBM: A Structured Latent Energy-Based Model for Blind Source Separation

DGX agent

arXiv:2604.17381v1 Announce Type: cross Abstract: This paper proposes StrEBM, a structured latent energy-based model for source-wise structured representation learning. The framework is motivated by a

researcharxiv-cs-lg
21 Apr 2026
Safety

SynopticBench: Evaluating Vision-Language Models on Generating Weather Forecast Discussions of the Future

DGX agent

arXiv:2604.16451v1 Announce Type: new Abstract: Recent advances in visual-language models (VLMs) have led to significant improvements in a plethora of complex multimodal tasks like image captioning, r

safetyarxiv-cs-cl
21 Apr 2026
Research

Test-Time Perturbation Learning with Delayed Feedback for Vision-Language-Action Models

DGX agent

arXiv:2604.18107v1 Announce Type: new Abstract: Vision-Language-Action models (VLAs) achieve remarkable performance in sequential decision-making but remain fragile to subtle environmental shifts, suc

researcharxiv-cs-cv
21 Apr 2026
Agents

The Global Neural World Model: Spatially Grounded Discrete Topologies for Action-Conditioned Planning

DGX agent

arXiv:2604.16585v1 Announce Type: new Abstract: We present the Global Neural World Model (GNWM), a self-stabilizing framework that achieves topological quantization through balanced continuous entropy

agentsarxiv-cs-lg
21 Apr 2026
Model Releases

TLoRA: Task-aware Low Rank Adaptation of Large Language Models

DGX agent

arXiv:2604.18124v1 Announce Type: new Abstract: Low-Rank Adaptation (LoRA) has become a widely adopted parameter-efficient fine-tuning method for large language models, with its effectiveness largely

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Towards Joint Quantization and Token Pruning of Vision-Language Models

DGX agent

arXiv:2604.17320v1 Announce Type: new Abstract: Deploying Vision-Language Models (VLMs) under aggressive low-bit inference remains challenging because inference cost is dominated by the long visual-to

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Unsupervised Discovery of Intermediate Phase Order in the Frustrated J_1-J_2 Heisenberg Model via Prometheus Framework

DGX agent

arXiv:2602.21468v4 Announce Type: replace-cross Abstract: The spin-1/2 J_1-J_2 Heisenberg model on the square lattice exhibits a debated intermediate phase between Neel antiferromagnetic and stripe or

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

When Text Hijacks Vision: Benchmarking and Mitigating Text Overlay-Induced Hallucination in Vision Language Models

DGX agent

arXiv:2604.17375v1 Announce Type: new Abstract: Recent advances in Vision-Language Models (VLMs) have substantially enhanced their ability across multimodal video understanding benchmarks spanning tem

model-releasesarxiv-cs-cv
21 Apr 2026
Applications

Applied Explainability for Large Language Models: A Comparative Study

DGX agent

arXiv:2604.15371v1 Announce Type: cross Abstract: Large language models (LLMs) achieve strong performance across many natural language processing tasks, yet their decision processes remain difficult t

applicationsarxiv-cs-ai
20 Apr 2026
Safety

AutoDrive-R^2: Incentivizing Reasoning and Self-Reflection Capacity for VLA Model in Autonomous Driving

DGX agent

arXiv:2509.01944v3 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models in autonomous driving systems have recently demonstrated transformative potential by integrating multimoda

safetyarxiv-cs-cv
20 Apr 2026
Model Releases

Automating Crash Diagram Generation Using Vision-Language Models: A Case Study on Multi-Lane Roundabouts

DGX agent

arXiv:2604.15332v1 Announce Type: cross Abstract: Crash diagrams are essential tools in transportation safety analysis, yet their manual preparation remains time-consuming and prone to human variabili

model-releasesarxiv-cs-ai
20 Apr 2026
Safety

Concept-wise Attention for Fine-grained Concept Bottleneck Models

DGX agent

arXiv:2604.15748v1 Announce Type: new Abstract: Recently impressive performance has been achieved in Concept Bottleneck Models (CBM) by utilizing the image-text alignment learned by a large pre-traine

safetyarxiv-cs-cv
20 Apr 2026
Model Releases

ConFu: Contemplate the Future for Better Speculative Sampling

DGX agent

arXiv:2603.08899v2 Announce Type: replace Abstract: Speculative decoding has emerged as a powerful approach to accelerate large language model (LLM) inference by employing lightweight draft models to

model-releasesarxiv-cs-cl
20 Apr 2026
Local Ai

DINOv3 Beats Specialized Detectors: A Simple Foundation Model Baseline for Image Forensics

DGX agent

arXiv:2604.16083v1 Announce Type: new Abstract: With the rapid advancement of deep generative models, realistic fake images have become increasingly accessible, yet existing localization methods rely

local-aiarxiv-cs-cv
20 Apr 2026
Model Releases

FETAL-GAUGE: A Benchmark for Assessing Vision-Language Models in Fetal Ultrasound

DGX agent

arXiv:2512.22278v2 Announce Type: replace Abstract: The growing demand for prenatal ultrasound imaging has intensified a global shortage of trained sonographers, creating barriers to essential fetal h

model-releasesarxiv-cs-cv
20 Apr 2026
Model Releases

HyperGVL: Benchmarking and Improving Large Vision-Language Models in Hypergraph Understanding and Reasoning

DGX agent

arXiv:2604.15648v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) consistently require new arenas to guide their expanding boundaries, yet their capabilities with hypergraphs remain

model-releasesarxiv-cs-cl
20 Apr 2026
Safety

Jailbreak Scaling Laws for Large Language Models: Polynomial-Exponential Crossover

DGX agent

arXiv:2603.11331v2 Announce Type: replace-cross Abstract: Adversarial attacks can reliably steer safety-aligned large language models toward unsafe behavior. Empirically, we find that strong adversari

safetyarxiv-cs-ai
20 Apr 2026
Model Releases

JumpLoRA: Sparse Adapters for Continual Learning in Large Language Models

DGX agent

arXiv:2604.16171v1 Announce Type: cross Abstract: Adapter-based methods have become a cost-effective approach to continual learning (CL) for Large Language Models (LLMs), by sequentially learning a lo

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

Reasoning-targeted Jailbreak Attacks on Large Reasoning Models via Semantic Triggers and Psychological Framing

DGX agent

arXiv:2604.15725v1 Announce Type: cross Abstract: Large Reasoning Models (LRMs) have demonstrated strong capabilities in generating step-by-step reasoning chains alongside final answers, enabling thei

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

RedBench: A Universal Dataset for Comprehensive Red Teaming of Large Language Models

DGX agent

arXiv:2601.03699v2 Announce Type: replace Abstract: As large language models (LLMs) become integral to safety-critical applications, ensuring their robustness against adversarial prompts is paramount.

model-releasesarxiv-cs-cl
20 Apr 2026
Model Releases

Seed1.8 Model Card: Towards Generalized Real-World Agency

DGX agent

arXiv:2603.20633v3 Announce Type: replace Abstract: We present Seed1.8, a foundation model aimed at generalized real-world agency: going beyond single-turn prediction to multi-turn interaction, tool u

model-releasesarxiv-cs-ai
20 Apr 2026
Research

SSMamba: A Self-Supervised Hybrid State Space Model for Pathological Image Classification

DGX agent

arXiv:2604.15711v1 Announce Type: cross Abstract: Pathological diagnosis is highly reliant on image analysis, where Regions of Interest (ROIs) serve as the primary basis for diagnostic evidence, while

researcharxiv-cs-ai
20 Apr 2026
Model Releases

Stargazer: A Scalable Model-Fitting Benchmark Environment for AI Agents under Astrophysical Constraints

DGX agent

arXiv:2604.15664v1 Announce Type: new Abstract: The rise of autonomous AI agents suggests that dynamic benchmark environments with built-in feedback on scientifically grounded tasks are needed to eval

model-releasesarxiv-cs-lg
20 Apr 2026
Safety

Towards Intrinsic Interpretability of Large Language Models:A Survey of Design Principles and Architectures

DGX agent

arXiv:2604.16042v1 Announce Type: cross Abstract: While Large Language Models (LLMs) have achieved strong performance across many NLP tasks, their opaque internal mechanisms hinder trustworthiness and

safetyarxiv-cs-ai
20 Apr 2026
Applications

Unveiling Stochasticity: Universal Multi-modal Probabilistic Modeling for Traffic Forecasting

DGX agent

arXiv:2604.16084v1 Announce Type: cross Abstract: Traffic forecasting is a challenging spatio-temporal modeling task and a critical component of urban transportation management. Current studies mainly

applicationsarxiv-cs-ai
20 Apr 2026
Model Releases

When Surfaces Lie: Exploiting Wrinkle-Induced Attention Shift to Attack Vision-Language Models

DGX agent

arXiv:2603.27759v3 Announce Type: replace Abstract: Visual-Language Models (VLMs) have demonstrated exceptional cross-modal understanding across various tasks, including zero-shot classification, imag

model-releasesarxiv-cs-cv
20 Apr 2026
Research

An Analysis of Regularization and Fokker-Planck Residuals in Diffusion Models for Image Generation

DGX agent

arXiv:2604.15171v1 Announce Type: new Abstract: Recent work has shown that diffusion models trained with the denoising score matching (DSM) objective often violate the Fokker--Planck (FP) equation tha

researcharxiv-cs-cv
17 Apr 2026
Agents

Empowerment Gain and Causal Model Construction: Children and adults are sensitive to controllability and variability in their causal interventions

DGX agent

arXiv:2512.08230v2 Announce Type: replace Abstract: Learning about the causal structure of the world is a fundamental problem for human cognition. Causal models and especially causal learning have pro

agentsarxiv-cs-ai
17 Apr 2026
Model Releases

EuropeMedQA Study Protocol: A Multilingual, Multimodal Medical Examination Dataset for Language Model Evaluation

DGX agent

arXiv:2604.14306v1 Announce Type: new Abstract: While Large Language Models (LLMs) have demonstrated high proficiency on English-centric medical examinations, their performance often declines when fac

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

GraphScout: Empowering Large Language Models with Intrinsic Exploration Ability for Agentic Graph Reasoning

DGX agent

arXiv:2603.01410v2 Announce Type: replace Abstract: Knowledge graphs provide structured and reliable information for many real-world applications, motivating increasing interest in combining large lan

model-releasesarxiv-cs-ai
17 Apr 2026
Model Releases

Improving Language Models with Intentional Analysis

DGX agent

arXiv:2502.04689v4 Announce Type: replace Abstract: Intent, a critical cognitive notion and mental state, is ubiquitous in human communication and problem-solving. Accurately understanding the underly

model-releasesarxiv-cs-cl
17 Apr 2026
Safety

Language of Thought Shapes Output Diversity in Large Language Models

DGX agent

arXiv:2601.11227v2 Announce Type: replace Abstract: Output diversity is crucial for Large Language Models as it underpins pluralism and creativity. In this work, we reveal that controlling the languag

safetyarxiv-cs-cl
17 Apr 2026
Safety

Multi-Persona Thinking for Bias Mitigation in Large Language Models

DGX agent

arXiv:2601.15488v2 Announce Type: replace Abstract: Large Language Models (LLMs) exhibit social biases, which can lead to harmful stereotypes and unfair outcomes. We propose extbf{Multi-Persona Thinki

safetyarxiv-cs-cl
17 Apr 2026
Tutorials

Pruning Long Chain-of-Thought of Large Reasoning Models via Small-Scale Preference Optimization

DGX agent

arXiv:2508.10164v2 Announce Type: replace Abstract: Recent advances in Large Reasoning Models (LRMs) have demonstrated strong performance on complex tasks through long Chain-of-Thought (CoT) reasoning

tutorialsarxiv-cs-ai
17 Apr 2026
Research

Random Matrix Theory for Deep Learning: Beyond Eigenvalues of Linear Models

DGX agent

arXiv:2506.13139v2 Announce Type: replace-cross Abstract: Modern Machine Learning (ML) and Deep Neural Networks (DNNs) often operate on high-dimensional data and rely on overparameterized models, wher

researcharxiv-cs-lg
17 Apr 2026
Safety

RaTA-Tool: Retrieval-based Tool Selection with Multimodal Large Language Models

DGX agent

arXiv:2604.14951v1 Announce Type: cross Abstract: Tool learning with foundation models aims to endow AI systems with the ability to invoke external resources -- such as APIs, computational utilities,

safetyarxiv-cs-cl
17 Apr 2026
Safety

RECOVER: Designing a Large Language Model-based Remote Patient Monitoring System for Postoperative Gastrointestinal Cancer Care

DGX agent

arXiv:2502.05740v2 Announce Type: replace-cross Abstract: Cancer surgery is a key treatment for gastrointestinal (GI) cancers, a group of cancers that account for more than 35% of cancer-related death

safetyarxiv-cs-ai
17 Apr 2026
Applications

Route to Rome Attack: Directing LLM Routers to Expensive Models via Adversarial Suffix Optimization

DGX agent

arXiv:2604.15022v1 Announce Type: cross Abstract: Cost-aware routing dynamically dispatches user queries to models of varying capability to balance performance and inference cost. However, the routing

applicationsarxiv-cs-cl
17 Apr 2026
← Previous
1…118119120121122…1030
Next →