AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

Content type
85,202Total entries
1Added by human
85,201Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,811 results
Safety

Representational Alignment Across Model Layers and Brain Regions with Multi-Level Optimal Transport

DGX agent

arXiv:2510.01706v2 Announce Type: replace-cross Abstract: Standard representational similarity methods align each layer of a network to its best match in another independently, producing asymmetric re

safetyarxiv-cs-ai
23 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

ThermoQA: A Three-Tier Benchmark for Evaluating Thermodynamic Reasoning in Large Language Models

DGX agent

arXiv:2604.19758v1 Announce Type: new Abstract: We present ThermoQA, a benchmark of 293 open-ended engineering thermodynamics problems in three tiers: property lookups (110 Q), component analysis (101

model-releasesarxiv-cs-ai
23 Apr 2026
Research

Tracing Relational Knowledge Recall in Large Language Models

DGX agent

arXiv:2604.19934v1 Announce Type: new Abstract: We study how large language models recall relational knowledge during text generation, with a focus on identifying latent representations suitable for r

researcharxiv-cs-cl
23 Apr 2026
Model Releases

Variance Is Not Importance: Structural Analysis of Transformer Compressibility Across Model Scales

DGX agent

arXiv:2604.20682v1 Announce Type: new Abstract: We present a systematic empirical study of transformer compression through over 40 experiments on GPT-2 (124M parameters) and Mistral 7B (7.24B paramete

model-releasesarxiv-cs-lg
23 Apr 2026
Safety

ARM: Advantage Reward Modeling for Long-Horizon Manipulation

DGX agent

arXiv:2604.03037v2 Announce Type: replace-cross Abstract: Long-horizon robotic manipulation remains challenging for reinforcement learning (RL) because sparse rewards provide limited guidance for cred

safetyarxiv-cs-ai
22 Apr 2026
Safety

Attention-based Multi-modal Deep Learning Model of Spatio-temporal Crop Yield Prediction with Satellite, Soil and Climate Data

DGX agent

arXiv:2604.19217v1 Announce Type: cross Abstract: Crop yield prediction is one of the most important challenge, which is crucial to world food security and policy-making decisions. The conventional fo

safetyarxiv-cs-ai
22 Apr 2026
Safety

Decomposed Trust: Privacy, Adversarial Robustness, Ethics, and Fairness in Low-Rank LLMs

DGX agent

arXiv:2511.22099v3 Announce Type: replace-cross Abstract: Large language models (LLMs) have driven major advances across domains, yet their massive size hinders deployment in resource-constrained sett

safetyarxiv-cs-ai
22 Apr 2026
Model Releases

Heterogeneity-Aware Personalized Federated Learning for Industrial Predictive Analytics

DGX agent

arXiv:2604.19451v1 Announce Type: new Abstract: Federated prognostics enable clients (e.g., companies, factories, and production lines) to collaboratively develop a failure time prediction model while

model-releasesarxiv-cs-lg
22 Apr 2026
Safety

Location Not Found: Exposing Implicit Local and Global Biases in Multilingual LLMs

DGX agent

arXiv:2604.19292v1 Announce Type: cross Abstract: Multilingual large language models (LLMs) have minimized the fluency gap between languages. This advancement, however, exposes models to the risk of b

safetyarxiv-cs-ai
22 Apr 2026
Model Releases

Lost in Translation: Do LVLM Judges Generalize Across Languages?

DGX agent

arXiv:2604.19405v1 Announce Type: new Abstract: Automatic evaluators such as reward models play a central role in the alignment and evaluation of large vision-language models (LVLMs). Despite their gr

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

SAGE-32B: Agentic Reasoning via Iterative Distillation

DGX agent

arXiv:2601.04237v2 Announce Type: replace Abstract: We demonstrate SAGE-32B, a 32 billion parameter language model that focuses on agentic reasoning and long range planning tasks. Unlike chat models t

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

The High Explosives and Affected Targets (HEAT) Dataset

DGX agent

arXiv:2604.18828v1 Announce Type: new Abstract: Artificial Intelligence (AI) surrogate models provide a computationally efficient alternative to full-physics simulations, but no public datasets curren

model-releasesarxiv-cs-lg
22 Apr 2026
Tutorials

A Survey of Reinforcement Learning for Large Language Models under Data Scarcity: Challenges and Solutions

DGX agent

arXiv:2604.17312v1 Announce Type: new Abstract: Reinforcement learning (RL) has emerged as a powerful post-training paradigm for enhancing the reasoning capabilities of large language models (LLMs). H

tutorialsarxiv-cs-lg
21 Apr 2026
Model Releases

AVRT: Audio-Visual Reasoning Transfer through Single-Modality Teachers

DGX agent

arXiv:2604.16617v1 Announce Type: new Abstract: Recent advances in reasoning models have shown remarkable progress in text-based domains, but transferring those capabilities to multimodal settings, e.

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Bielik Guard: Efficient Polish Language Safety Classifiers for LLM Content Moderation

DGX agent

arXiv:2602.07954v4 Announce Type: replace Abstract: As Large Language Models (LLMs) become increasingly deployed in Polish language applications, the need for efficient and accurate content safety cla

model-releasesarxiv-cs-cl
21 Apr 2026
Agents

CLAG: Adaptive Memory Organization via Agent-Driven Clustering for Small Language Model Agents

DGX agent

arXiv:2603.15421v2 Announce Type: replace Abstract: Large language model agents heavily rely on external memory to support knowledge reuse and complex reasoning tasks. Yet most memory systems store ex

agentsarxiv-cs-cl
21 Apr 2026
Model Releases

Cognitive Chain-of-Thought (CoCoT): Structured Multimodal Reasoning about Social Situations

DGX agent

arXiv:2507.20409v2 Announce Type: replace Abstract: Chain-of-Thought (CoT) prompting helps models think step by step. But naive CoT breaks down in visually grounded social tasks, where models must per

model-releasesarxiv-cs-cl
21 Apr 2026
Applications

Compositional Steering of Large Language Models with Steering Tokens

DGX agent

arXiv:2601.05062v2 Announce Type: replace Abstract: Deploying LLMs in real-world applications requires controllable output that satisfies multiple desiderata at the same time. While existing work exte

applicationsarxiv-cs-cl
21 Apr 2026
Research

D-Prism: Differentiable Primitives for Structured Dynamic Modeling

DGX agent

arXiv:2604.17082v1 Announce Type: new Abstract: Capturing both geometry and rigid motion for structured dynamic objects, like multi-part assemblies or jointed mechanisms, remains a key challenge. Exis

researcharxiv-cs-cv
21 Apr 2026
Safety

DeepThinkVLA: Enhancing Reasoning Capability of Vision-Language-Action Models

DGX agent

arXiv:2511.15669v2 Announce Type: replace Abstract: Does Chain-of-Thought (CoT) reasoning genuinely improve Vision-Language-Action (VLA) models, or does it merely add overhead? Existing CoT-VLA system

safetyarxiv-cs-lg
21 Apr 2026
Research

Defragmenting Language Models: An Interpretability-based Approach for Vocabulary Expansion

DGX agent

arXiv:2604.16656v1 Announce Type: new Abstract: All languages are equal; when it comes to tokenization, some are more equal than others. Tokens are the hidden currency that dictate the cost and latenc

researcharxiv-cs-cl
21 Apr 2026
Research

emg2speech: Synthesizing speech from electromyography using self-supervised speech models

DGX agent

arXiv:2510.23969v2 Announce Type: replace-cross Abstract: We present a neuromuscular speech interface that translates electromyographic (EMG) signals recorded from orofacial muscles during speech arti

researcharxiv-cs-cl
21 Apr 2026
Agents

From Clinical Intent to Clinical Model: An Autonomous Coding-Agent Framework for Clinician-driven AI Development

DGX agent

arXiv:2604.17110v1 Announce Type: new Abstract: Clinical AI development has traditionally followed a collaborative paradigm that depends on close interaction between clinicians and specialized AI team

agentsarxiv-cs-cv
21 Apr 2026
Model Releases

From Handwriting to Structured Data: Benchmarking AI Digitisation of Handwritten Forms

DGX agent

arXiv:2604.16504v1 Announce Type: new Abstract: Manual digitisation of structured handwritten documents is slow and costly. We benchmark 17 leading frontier multi-modal large language models and open-

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

How Robustly do LLMs Understand Execution Semantics?

DGX agent

arXiv:2604.16320v1 Announce Type: cross Abstract: LLMs demonstrate remarkable reasoning capabilities, yet whether they utilize internal world models or rely on sophisticated pattern matching remains o

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Language Models Don't Know What You Want: Evaluating Personalization in Deep Research Needs Real Users

DGX agent

arXiv:2603.16120v2 Announce Type: replace Abstract: Deep Research (DR) systems help researchers cope with ballooning publishing counts. Such tools synthesize scientific papers to answer research queri

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Latent Preference Modeling for Cross-Session Personalized Tool Calling

DGX agent

arXiv:2604.17886v1 Announce Type: new Abstract: Users often omit essential details in their requests to LLM-based agents, resulting in under-specified inputs for tool use. This poses a fundamental cha

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Long-Text-to-Image Generation via Compositional Prompt Decomposition

DGX agent

arXiv:2604.18258v1 Announce Type: new Abstract: While modern text-to-image (T2I) models excel at generating images from intricate prompts, they struggle to capture the key details when the inputs are

model-releasesarxiv-cs-cv
21 Apr 2026
Research

LoReC: Rethinking Large Language Models for Graph Data Analysis

DGX agent

arXiv:2604.17897v1 Announce Type: new Abstract: The advent of Large Language Models (LLMs) has fundamentally reshaped the way we interact with graphs, giving rise to a new paradigm called GraphLLM. As

researcharxiv-cs-lg
21 Apr 2026
Research

Modeling, Control and Self-sensing of Dielectric Elastomer Soft Actuators: A Review

DGX agent

arXiv:2604.17199v1 Announce Type: new Abstract: Dielectric elastomer actuators (DEAs) have garnered extensive attention especially in soft robotic applications over the past few decades owing to the a

researcharxiv-cs-ro
21 Apr 2026
Model Releases

Modeling Higher-Order Brain Interactions via a Multi-View Information Bottleneck Framework for fMRI-based Psychiatric Diagnosis

DGX agent

arXiv:2604.17713v1 Announce Type: new Abstract: Resting-state functional magnetic resonance imaging (fMRI) has emerged as a cornerstone for psychiatric diagnosis, yet most approaches rely on pairwise

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Modelling Gas-Phase Reaction Kinetics with Guided Particle Diffusion Sampling

DGX agent

arXiv:2604.16461v1 Announce Type: cross Abstract: Physics-guided sampling with diffusion priors has recently shown strong performance in solving complex systems of partial differential equations (PDEs

model-releasesarxiv-cs-lg
21 Apr 2026
Safety

Operationalizing Fairness in Text-to-Image Models: A Survey of Bias, Fairness Audits and Mitigation Strategies

DGX agent

arXiv:2604.16516v1 Announce Type: new Abstract: Text-to-Image (T2I) generation models have been widely adopted across various industries, yet are criticized for frequently exhibiting societal stereoty

safetyarxiv-cs-cv
21 Apr 2026
Model Releases

PiCa: Parameter-Efficient Fine-Tuning with Column Space Projection

DGX agent

arXiv:2505.20211v3 Announce Type: replace Abstract: Fine-tuning large foundation models is essential for building expert models tailored to specialized tasks and domains, but fully updating billions o

model-releasesarxiv-cs-lg
21 Apr 2026
Safety

Safer Trajectory Planning with CBF-guided Diffusion Model for Unmanned Aerial Vehicles

DGX agent

arXiv:2604.17527v1 Announce Type: new Abstract: Safe and agile trajectory planning is essential for autonomous systems, especially during complex aerobatic maneuvers. Motivated by the recent success o

safetyarxiv-cs-ro
21 Apr 2026
Applications

Shifting the Gradient: Understanding How Defensive Training Methods Protect Language Model Integrity

DGX agent

arXiv:2604.16423v1 Announce Type: new Abstract: Defensive training methods such as positive preventative steering (PPS) and inoculation prompting (IP) offer surprising results through seemingly simila

applicationsarxiv-cs-lg
21 Apr 2026
Model Releases

SinkRouter: Sink-Aware Routing for Efficient Long-Context Decoding in Large Language and Multimodal Models

DGX agent

arXiv:2604.16883v1 Announce Type: new Abstract: In long-context decoding for LLMs and LMMs, attention becomes increasingly memory-bound because each decoding step must load a large amount of KV-cache

model-releasesarxiv-cs-lg
21 Apr 2026
Research

Spotlights and Blindspots: Evaluation Machine-Generated Text Detection

DGX agent

arXiv:2604.16607v1 Announce Type: new Abstract: With the rise of generative language models, machine-generated text detection has become a critical challenge. A wide variety of models is available, bu

researcharxiv-cs-cl
21 Apr 2026
Research

Target Parameterization in Diffusion Models for Nonlinear Spatiotemporal System Identification

DGX agent

arXiv:2604.17566v1 Announce Type: cross Abstract: Machine learning is becoming increasingly important for nonlinear system identification, including dynamical systems with spatially distributed output

researcharxiv-cs-lg
21 Apr 2026
Model Releases

Too Correct to Learn: Reinforcement Learning on Saturated Reasoning Data

DGX agent

arXiv:2604.18493v1 Announce Type: new Abstract: Reinforcement Learning (RL) enhances LLM reasoning, yet a paradox emerges as models scale: strong base models saturate standard benchmarks (e.g., MATH),

model-releasesarxiv-cs-lg
21 Apr 2026
Research

Towards Real-Time ECG and EMG Modeling on mu NPUs

DGX agent

arXiv:2604.18067v1 Announce Type: new Abstract: The miniaturisation of neural processing units (NPUs) and other low-power accelerators has enabled their integration into microcontroller-scale wearable

researcharxiv-cs-lg
21 Apr 2026
Research

Curing Miracle Steps in LLM Mathematical Reasoning with Rubric Rewards

DGX agent

arXiv:2510.07774v3 Announce Type: replace Abstract: In this paper, we observe that current models are susceptible to reward hacking, leading to a substantial overestimation of a model's reasoning abil

researcharxiv-cs-cl
20 Apr 2026
Model Releases

DriveLaW:Unifying Planning and Video Generation in a Latent Driving World

DGX agent

arXiv:2512.23421v3 Announce Type: replace Abstract: World models have become crucial for autonomous driving, as they learn how scenarios evolve over time to address the long-tail challenges of the rea

model-releasesarxiv-cs-cv
20 Apr 2026
Research

Early Detection of Acute Myeloid Leukemia (AML) Using YOLOv12 Deep Learning Model

DGX agent

arXiv:2604.16082v1 Announce Type: cross Abstract: Acute Myeloid Leukemia (AML) is one of the most life-threatening type of blood cancers, and its accurate classification is considered and remains a ch

researcharxiv-cs-ai
20 Apr 2026
Safety

Follow the Flow: On Information Flow Across Textual Tokens in Text-to-Image Models

DGX agent

arXiv:2504.01137v3 Announce Type: replace Abstract: Text-to-image generation models suffer from alignment problems, where generated images fail to accurately capture the objects and relations in the t

safetyarxiv-cs-cl
20 Apr 2026
Safety

Prototype-Grounded Concept Models for Verifiable Concept Alignment

DGX agent

arXiv:2604.16076v1 Announce Type: cross Abstract: Concept Bottleneck Models (CBMs) aim to improve interpretability in Deep Learning by structuring predictions through human-understandable concepts, bu

safetyarxiv-cs-ai
20 Apr 2026
Model Releases

The Spectral Geometry of Thought: Phase Transitions, Instruction Reversal, Token-Level Dynamics, and Perfect Correctness Prediction in How Transformers Reason

DGX agent

arXiv:2604.15350v1 Announce Type: new Abstract: We discover that large language models exhibit spectral phase transitions in their hidden activation spaces when engaging in reasoning versus factual re

model-releasesarxiv-cs-lg
20 Apr 2026
Local Ai

Where Do Vision-Language Models Fail? World Scale Analysis for Image Geolocalization

DGX agent

arXiv:2604.16248v1 Announce Type: new Abstract: Image geolocalization has traditionally been addressed through retrieval-based place recognition or geometry-based visual localization pipelines. Recent

local-aiarxiv-cs-cv
20 Apr 2026
← Previous
1…195196197198199…1038
Next →