AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,338
  • Agents7,718
  • Applications5,508
  • Concepts5
  • Hardware1,906
  • Industry6,195
  • Local Ai5,052
  • Model Releases24,549
  • Research20,616
  • Safety13,637
  • Syntheses17
  • Tools1,678
  • Tutorials3,457

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,338
  • Agents7,718
  • Applications5,508
  • Concepts5
  • Hardware1,906
  • Industry6,195
  • Local Ai5,052
  • Model Releases24,549
  • Research20,616
  • Safety13,637
  • Syntheses17
  • Tools1,678
  • Tutorials3,457

Source
HumanDGX agent

Content type
90,338Total entries
1Added by human
90,337Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
53,226 results
Model Releases

Robust Fair Disease Diagnosis in CT Images

DGX agent

arXiv:2604.09710v1 Announce Type: new Abstract: Automated diagnosis from chest CT has improved considerably with deep learning, but models trained on skewed datasets tend to perform unevenly across pa

model-releasesarxiv-cs-cv
14 Apr 2026
Local Ai

RPA-Check: A Multi-Stage Automated Framework for Evaluating Dynamic LLM-based Role-Playing Agents

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2604.11655v1 Announce Type: cross Abstract: The rapid adoption of Large Language Models (LLMs) in interactive systems has enabled the creation of dynamic, open-ended Role-Playing Agents (RPAs).

local-aiarxiv-cs-ai
14 Apr 2026
Model Releases

SCMAPR: Self-Correcting Multi-Agent Prompt Refinement for Complex-Scenario Text-to-Video Generation

DGX agent

arXiv:2604.05489v3 Announce Type: replace Abstract: Text-to-Video (T2V) generation has benefited from recent advances in diffusion models, yet current systems still struggle under complex scenarios, w

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

SecureVibeBench: Evaluating Secure Coding Capabilities of Code Agents with Realistic Vulnerability Scenarios

DGX agent

arXiv:2509.22097v3 Announce Type: replace-cross Abstract: Large language model-powered code agents are rapidly transforming software engineering, yet the security risks of their generated code have be

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Semantic-Geometric Dual Compression: Training-Free Visual Token Reduction for Ultra-High-Resolution Remote Sensing Understanding

DGX agent

arXiv:2604.11122v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have demonstrated immense potential in Earth observation. However, the massive visual tokens generated when p

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

Sense Less, Infer More: Agentic Multimodal Transformers for Edge Medical Intelligence

DGX agent

arXiv:2604.10404v1 Announce Type: cross Abstract: Edge-based multimodal medical monitoring requires models that balance diagnostic accuracy with severe energy constraints. Continuous acquisition of EC

safetyarxiv-cs-lg
14 Apr 2026
Model Releases

Solving Physics Olympiad via Reinforcement Learning on Physics Simulators

DGX agent

arXiv:2604.11805v1 Announce Type: cross Abstract: We have witnessed remarkable advances in LLM reasoning capabilities with the advent of DeepSeek-R1. However, much of this progress has been fueled by

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

Steered LLM Activations are Non-Surjective

DGX agent

arXiv:2604.09839v1 Announce Type: new Abstract: Activation steering is a popular white-box control technique that modifies model activations to elicit an abstract change in output behavior. It has als

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

Switch-JustDance: Benchmarking Whole Body Motion Tracking Controllers Using a Commercial Console Game

DGX agent

arXiv:2511.17925v3 Announce Type: replace-cross Abstract: Recent advances in whole-body robot control have enabled humanoid and legged robots to perform increasingly agile and coordinated motions. How

model-releasesarxiv-cs-cv
14 Apr 2026
Research

TARAC: Mitigating Hallucination in LVLMs via Temporal Attention Real-time Accumulative Connection

DGX agent

arXiv:2504.04099v2 Announce Type: replace-cross Abstract: Large Vision-Language Models have demonstrated remarkable capabilities, yet they suffer from hallucinations that limit practical deployment. W

researcharxiv-cs-ai
14 Apr 2026
Research

Teaching Robots to Interpret Social Interactions through Lexically-guided Dynamic Graph Learning

DGX agent

arXiv:2604.10895v1 Announce Type: cross Abstract: For a robot to be called socially intelligent, it must be able to infer users internal states from their current behaviour, predict the users future b

researcharxiv-cs-ro
14 Apr 2026
Model Releases

The Blind Spot of Agent Safety: How Benign User Instructions Expose Critical Vulnerabilities in Computer-Use Agents

DGX agent

arXiv:2604.10577v1 Announce Type: cross Abstract: Computer-use agents (CUAs) can now autonomously complete complex tasks in real digital environments, but when misled, they can also be used to automat

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Think Parallax: Solving Multi-Hop Problems via Multi-View Knowledge-Graph-Based Retrieval-Augmented Generation

DGX agent

arXiv:2510.15552v3 Announce Type: replace-cross Abstract: Large language models (LLMs) still struggle with multi-hop reasoning over knowledge-graphs (KGs), and we identify a previously overlooked stru

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Topo-ADV: Generating Topology-Driven Imperceptible Adversarial Point Clouds

DGX agent

arXiv:2604.09879v1 Announce Type: new Abstract: Deep neural networks for 3D point cloud understanding have achieved remarkable success in object classification and recognition, yet recent work shows t

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Tracing the Roots: A Multi-Agent Framework for Uncovering Data Lineage in Post-Training LLMs

DGX agent

arXiv:2604.10480v1 Announce Type: new Abstract: Post-training data plays a pivotal role in shaping the capabilities of Large Language Models (LLMs), yet datasets are often treated as isolated artifact

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Training-Free Object-Background Compositional T2I via Dynamic Spatial Guidance and Multi-Path Pruning

DGX agent

arXiv:2604.09850v1 Announce Type: new Abstract: Existing text-to-image diffusion models, while excelling at subject synthesis, exhibit a persistent foreground bias that treats the background as a pass

model-releasesarxiv-cs-cv
14 Apr 2026
Safety

Weird Generalization is Weirdly Brittle

DGX agent

arXiv:2604.10022v1 Announce Type: new Abstract: Weird generalization is a phenomenon in which models fine-tuned on data from a narrow domain (e.g. insecure code) develop surprising traits that manifes

safetyarxiv-cs-cl
14 Apr 2026
Model Releases

What and Where to Adapt: Structure-Semantics Co-Tuning for Machine Vision Compression via Synergistic Adapters

DGX agent

arXiv:2604.10017v1 Announce Type: new Abstract: Parameter-efficient fine-tuning of pre-trained codecs is a promising direction in image compression for human and machine vision. While most existing wo

model-releasesarxiv-cs-cv
14 Apr 2026
Safety

What Factors Affect LLMs and RLLMs in Financial Question Answering?

DGX agent

arXiv:2507.08339v4 Announce Type: replace Abstract: Recently, large language models (LLMs) and reasoning large language models (RLLMs) have gained considerable attention from many researchers. RLLMs e

safetyarxiv-cs-cl
14 Apr 2026
Model Releases

Who Wrote This Line? Evaluating the Detection of LLM-Generated Classical Chinese Poetry

DGX agent

arXiv:2604.10101v1 Announce Type: new Abstract: The rapid development of large language models (LLMs) has extended text generation tasks into the literary domain. However, AI-generated literary creati

model-releasesarxiv-cs-cl
14 Apr 2026
Research

XD-MAP: Cross-Modal Domain Adaptation via Semantic Parametric Maps for Scalable Training Data Generation

DGX agent

arXiv:2601.14477v2 Announce Type: replace-cross Abstract: Until open-world foundation models match the performance of specialized approaches, deep learning systems remain dependent on task- and sensor

researcharxiv-cs-ai
14 Apr 2026
Model Releases

4D-RGPT: Toward Region-level 4D Understanding via Perceptual Distillation

DGX agent

arXiv:2512.17012v3 Announce Type: replace Abstract: Despite advances in Multimodal LLMs (MLLMs), their ability to reason over 3D structures and temporal dynamics remains limited, constrained by weak 4

model-releasesarxiv-cs-cv
13 Apr 2026
Research

Across the Levels of Analysis: Explaining Predictive Processing in Humans Requires More Than Machine-Estimated Probabilities

DGX agent

arXiv:2604.09466v1 Announce Type: new Abstract: Under the lens of Marr's levels of analysis, we critique and extend two claims about language models (LMs) and language processing: first, that predicti

researcharxiv-cs-cl
13 Apr 2026
Model Releases

AgentCE-Bench: Agent Configurable Evaluation with Scalable Horizons and Controllable Difficulty under Lightweight Environments

DGX agent

arXiv:2604.06111v2 Announce Type: replace Abstract: Existing Agent benchmarks suffer from two critical limitations: high environment interaction overhead (up to 41% of total evaluation time) and imbal

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Automated Instruction Revision (AIR): A Structured Comparison of Task Adaptation Strategies for LLM

DGX agent

arXiv:2604.09418v1 Announce Type: new Abstract: This paper studies Automated Instruction Revision (AIR), a rule-induction-based method for adapting large language models (LLMs) to downstream tasks usi

model-releasesarxiv-cs-cl
13 Apr 2026
Safety

Better Eyes, Better Thoughts: Why Vision Chain-of-Thought Fails in Medicine

DGX agent

arXiv:2603.06665v2 Announce Type: replace-cross Abstract: Large vision-language models (VLMs) often benefit from chain-of-thought (CoT) prompting in general domains, yet its efficacy in medical vision

safetyarxiv-cs-ai
13 Apr 2026
Model Releases

Bharat Scene Text: A Novel Comprehensive Dataset and Benchmark for Indian Language Scene Text Understanding

DGX agent

arXiv:2511.23071v2 Announce Type: replace-cross Abstract: Reading scene text, that is, text appearing in images, has numerous application areas, including assistive technology, search, and e-commerce.

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Boosting Brain-inspired Path Integration Efficiency via Learning-based Replication of Continuous Attractor Neurodynamics

DGX agent

arXiv:2511.17687v2 Announce Type: replace Abstract: The brain's Path Integration (PI) mechanism offers substantial guidance and inspiration for Brain-Inspired Navigation (BIN). However, the PI capabil

model-releasesarxiv-cs-lg
13 Apr 2026
Model Releases

CAD 100K: A Comprehensive Multi-Task Dataset for Car Related Visual Anomaly Detection

DGX agent

arXiv:2604.09023v1 Announce Type: new Abstract: Multi-task visual anomaly detection is critical for car-related manufacturing quality assessment. However, existing methods remain task-specific, hinder

model-releasesarxiv-cs-cv
13 Apr 2026
Model Releases

DeFakeQ: Enabling Real-Time Deepfake Detection on Edge Devices via Adaptive Bidirectional Quantization

DGX agent

arXiv:2604.08847v1 Announce Type: new Abstract: Deepfake detection has become a fundamental component of modern media forensics. Despite significant progress in detection accuracy, most existing metho

model-releasesarxiv-cs-cv
13 Apr 2026
Safety

Dictionary-Aligned Concept Control for Safeguarding Multimodal LLMs

DGX agent

arXiv:2604.08846v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have been shown to be vulnerable to malicious queries that can elicit unsafe responses. Recent work uses prom

safetyarxiv-cs-ai
13 Apr 2026
Model Releases

Fine-Grained Action Segmentation for Renorrhaphy in Robot-Assisted Partial Nephrectomy

DGX agent

arXiv:2604.09051v1 Announce Type: new Abstract: Fine-grained action segmentation during renorrhaphy in robot-assisted partial nephrectomy requires frame-level recognition of visually similar suturing

model-releasesarxiv-cs-cv
13 Apr 2026
Model Releases

FIRE-CIR: Fine-grained Reasoning for Composed Fashion Image Retrieval

DGX agent

arXiv:2604.09114v1 Announce Type: new Abstract: Composed image retrieval (CIR) aims to retrieve a target image that depicts a reference image modified by a textual description. While recent vision-lan

model-releasesarxiv-cs-cv
13 Apr 2026
Safety

GRM: Utility-Aware Jailbreak Attacks on Audio LLMs via Gradient-Ratio Masking

DGX agent

arXiv:2604.09222v1 Announce Type: cross Abstract: Audio large language models (ALLMs) enable rich speech-text interaction, but they also introduce jailbreak vulnerabilities in the audio modality. Exis

safetyarxiv-cs-ai
13 Apr 2026
Research

How Noise Benefits AI-generated Image Detection

DGX agent

arXiv:2511.16136v2 Announce Type: replace Abstract: The rapid advancement of generative models has made real and synthetic images increasingly indistinguishable. Although extensive efforts have been d

researcharxiv-cs-cv
13 Apr 2026
Research

Is More Data Worth the Cost? Dataset Scaling Laws in a Tiny Attention-Only Decoder

DGX agent

arXiv:2604.09389v1 Announce Type: cross Abstract: Training Transformer language models is expensive, as performance typically improves with increasing dataset size and computational budget. Although s

researcharxiv-cs-cl
13 Apr 2026
Applications

LLM-Rosetta: A Hub-and-Spoke Intermediate Representation for Cross-Provider LLM API Translation

DGX agent

arXiv:2604.09360v1 Announce Type: cross Abstract: The rapid proliferation of Large Language Model (LLM) providers--each exposing proprietary API formats--has created a fragmented ecosystem where appli

applicationsarxiv-cs-ai
13 Apr 2026
Model Releases

MARINER: A 3E-Driven Benchmark for Fine-Grained Perception and Complex Reasoning in Open-Water Environments

DGX agent

arXiv:2604.08615v1 Announce Type: cross Abstract: Fine-grained visual understanding and high-level reasoning in real-world open-water environments remain under-explored due to the lack of dedicated be

model-releasesarxiv-cs-ai
13 Apr 2026
Research

Measurement-Consistent Langevin Corrector for Stabilizing Latent Diffusion Inverse Problem Solvers

DGX agent

arXiv:2601.04791v3 Announce Type: replace Abstract: While latent diffusion models (LDMs) have emerged as powerful priors for inverse problems, existing LDM-based solvers frequently suffer from instabi

researcharxiv-cs-cv
13 Apr 2026
Safety

Mechanisms of Introspective Awareness

DGX agent

arXiv:2603.21396v2 Announce Type: replace Abstract: Recent work has shown that LLMs can sometimes detect when steering vectors are injected into their residual stream and identify the injected concept

safetyarxiv-cs-lg
13 Apr 2026
Model Releases

MolPaQ: Modular Quantum-Classical Patch Learning for Interpretable Molecular Generation

DGX agent

arXiv:2604.08575v1 Announce Type: cross Abstract: Molecular generative models must jointly ensure validity, diversity, and property control, yet existing approaches typically trade off among these obj

model-releasesarxiv-cs-ai
13 Apr 2026
Safety

Mosaic: Multimodal Jailbreak against Closed-Source VLMs via Multi-View Ensemble Optimization

DGX agent

arXiv:2604.09253v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) are powerful but remain vulnerable to multimodal jailbreak attacks. Existing attacks mainly rely on either explicit visu

safetyarxiv-cs-ai
13 Apr 2026
Agents

MT-OSC: Path for LLMs that Get Lost in Multi-Turn Conversation

DGX agent

arXiv:2604.08782v1 Announce Type: new Abstract: Large language models (LLMs) suffer significant performance degradation when user instructions and context are distributed over multiple conversational

agentsarxiv-cs-cl
13 Apr 2026
Model Releases

Multivariate Time Series Anomaly Detection via Dual-Branch Reconstruction and Autoregressive Flow-based Residual Density Estimation

DGX agent

arXiv:2604.08582v1 Announce Type: cross Abstract: Multivariate Time Series Anomaly Detection (MTSAD) is critical for real-world monitoring scenarios such as industrial control and aerospace systems. M

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Natural Riemannian gradient for learning functional tensor networks

DGX agent

arXiv:2604.09263v1 Announce Type: cross Abstract: We consider machine learning tasks with low-rank functional tree tensor networks (TTN) as the learning model. While in the case of least-squares regre

model-releasesarxiv-cs-lg
13 Apr 2026
Safety

Neural Distribution Prior for LiDAR Out-of-Distribution Detection

DGX agent

arXiv:2604.09232v1 Announce Type: cross Abstract: LiDAR-based perception is critical for autonomous driving due to its robustness to poor lighting and visibility conditions. Yet, current models operat

safetyarxiv-cs-ai
13 Apr 2026
Local Ai

Neurons Speak in Ranges: Breaking Free from Discrete Neuronal Attribution

DGX agent

arXiv:2502.06809v3 Announce Type: replace-cross Abstract: Pervasive polysemanticity in large language models (LLMs) undermines discrete neuron-concept attribution, posing a significant challenge for m

local-aiarxiv-cs-ai
13 Apr 2026
Local Ai

Offline-First LLM Architecture for Adaptive Learning in Low-Connectivity Environments

DGX agent

arXiv:2603.03339v5 Announce Type: replace-cross Abstract: Artificial intelligence (AI) and large language models (LLMs) are transforming educational technology by enabling conversational tutoring, per

local-aiarxiv-cs-cl
13 Apr 2026
← Previous
1…462463464465466…1109
Next →