AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,435 results
Model Releases

RepIt: Steering Language Models with Concept-Specific Refusal Vectors

DGX agent

arXiv:2509.13281v5 Announce Type: replace Abstract: Current safety evaluations of language models rely on benchmark-based assessments that may miss localized vulnerabilities. We present RepIt, a simpl

model-releasesarxiv-cs-ai
22 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

UniT: Toward a Unified Physical Language for Human-to-Humanoid Policy Learning and World Modeling

DGX agent

arXiv:2604.19734v1 Announce Type: cross Abstract: Scaling humanoid foundation models is bottlenecked by the scarcity of robotic data. While massive egocentric human data offers a scalable alternative,

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

VLA Foundry: A Unified Framework for Training Vision-Language-Action Models

DGX agent

arXiv:2604.19728v1 Announce Type: cross Abstract: We present VLA Foundry, an open-source framework that unifies LLM, VLM, and VLA training in a single codebase. Most open-source VLA efforts specialize

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

A Transformer and Prototype-based Interpretable Model for Contextual Sarcasm Detection

DGX agent

arXiv:2503.11838v2 Announce Type: replace Abstract: Sarcasm detection, with its figurative nature, poses unique challenges for affective systems designed to perform sentiment analysis. While these sys

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Agentic Large Language Models for Training-Free Neuro-Radiological Image Analysis

DGX agent

arXiv:2604.16729v1 Announce Type: new Abstract: State-of-the-art large language models (LLMs) show high performance in general visual question answering. However, a fundamental limitation remains: cur

model-releasesarxiv-cs-cv
21 Apr 2026
Local Ai

Aligning Language Models with Real-time Knowledge Editing

DGX agent

arXiv:2508.01302v3 Announce Type: replace Abstract: Knowledge editing aims to modify outdated knowledge in language models efficiently while retaining their original capabilities. Mainstream datasets

local-aiarxiv-cs-cl
21 Apr 2026
Model Releases

AnchorMem: Anchored Facts with Associative Contexts for Building Memory in Large Language Models

DGX agent

arXiv:2604.17377v1 Announce Type: new Abstract: While large language models have achieved remarkable performance in complex tasks, they still need a memory system to utilize historical experience in l

model-releasesarxiv-cs-cl
21 Apr 2026
Research

BARD: Bridging AutoRegressive and Diffusion Vision-Language Models Via Highly Efficient Progressive Block Merging and Stage-Wise Distillation

DGX agent

arXiv:2604.16514v1 Announce Type: new Abstract: Autoregressive vision-language models (VLMs) deliver strong multimodal capability, but their token-by-token decoding imposes a fundamental inference bot

researcharxiv-cs-cv
21 Apr 2026
Model Releases

BOP-ASK: Object-Interaction Reasoning for Vision-Language Models

DGX agent

arXiv:2511.16857v3 Announce Type: replace Abstract: Vision Language Models (VLMs) have achieved impressive performance on spatial reasoning benchmarks, yet these evaluations mask critical weaknesses i

model-releasesarxiv-cs-cv
21 Apr 2026
Tutorials

ControlAudio: Tackling Text-Guided, Timing-Indicated and Intelligible Audio Generation via Progressive Diffusion Modeling

DGX agent

arXiv:2510.08878v3 Announce Type: replace-cross Abstract: Text-to-audio (TTA) generation with fine-grained control signals, e.g., precise timing control or intelligible speech content, has been explor

tutorialsarxiv-cs-cl
21 Apr 2026
Research

Emergent Structured Representations Support Flexible In-Context Inference in Large Language Models

DGX agent

arXiv:2602.07794v3 Announce Type: replace Abstract: Large language models (LLMs) exhibit emergent behaviors suggestive of human-like reasoning. While recent work has identified structured conceptual r

researcharxiv-cs-cl
21 Apr 2026
Model Releases

Enhancing Continual Learning of Vision-Language Models via Dynamic Prefix Weighting

DGX agent

arXiv:2604.18075v1 Announce Type: new Abstract: We investigate recently introduced domain-class incremental learning scenarios for vision-language models (VLMs). Recent works address this challenge us

model-releasesarxiv-cs-cv
21 Apr 2026
Research

Evalet: Evaluating Large Language Models through Functional Fragmentation

DGX agent

arXiv:2509.11206v4 Announce Type: replace-cross Abstract: Practitioners increasingly rely on Large Language Models (LLMs) to evaluate generative AI outputs through 'LLM-as-a-Judge' approaches. However

researcharxiv-cs-cl
21 Apr 2026
Research

Follow the Path: Reasoning over Knowledge Graph Paths to Improve Large Language Model Factuality

DGX agent

arXiv:2505.11140v3 Announce Type: replace Abstract: We introduce fs1, a simple yet effective method that improves the factuality of reasoning traces by collecting them from large reasoning models and

researcharxiv-cs-cl
21 Apr 2026
Applications

From Static Inference to Dynamic Interaction: A Survey of Streaming Large Language Models

DGX agent

arXiv:2603.04592v3 Announce Type: replace Abstract: Standard Large Language Models (LLMs) are predominantly designed for static inference with pre-defined inputs, which limits their applicability in d

applicationsarxiv-cs-cl
21 Apr 2026
Research

HiPrune: Hierarchical Attention for Efficient Token Pruning in Vision-Language Models

DGX agent

arXiv:2508.00553v3 Announce Type: replace Abstract: Vision-Language Models (VLMs) encode images and videos into abundant tokens, which contain substantial redundancy and computation cost. While visual

researcharxiv-cs-cv
21 Apr 2026
Model Releases

HQA-VLAttack: Towards High Quality Adversarial Attack on Vision-Language Pre-Trained Models

DGX agent

arXiv:2604.16499v1 Announce Type: new Abstract: Black-box adversarial attack on vision-language pre-trained models is a practical and challenging task, as text and image perturbations need to be consi

model-releasesarxiv-cs-cv
21 Apr 2026
Applications

Large Language Models Are Bad Dice Players: LLMs Struggle to Generate Random Numbers from Statistical Distributions

DGX agent

arXiv:2601.05414v2 Announce Type: replace Abstract: As large language models (LLMs) transition from chat interfaces to integral components of stochastic pipelines and systems approaching general intel

applicationsarxiv-cs-cl
21 Apr 2026
Research

Linear-Time and Constant-Memory Text Embeddings Based on Recurrent Language Models

DGX agent

arXiv:2604.18199v1 Announce Type: new Abstract: Transformer-based embedding models suffer from quadratic computational and linear memory complexity, limiting their utility for long sequences. We propo

researcharxiv-cs-cl
21 Apr 2026
Safety

Mammo-FM: Breast-specific foundational model for Integrated Mammographic Diagnosis, Prognosis, and Reporting

DGX agent

arXiv:2512.00198v2 Announce Type: replace Abstract: Breast cancer is one of the leading causes of death among women worldwide. We introduce Mammo-FM, the first foundation model specifically for mammog

safetyarxiv-cs-cv
21 Apr 2026
Model Releases

Measuring Social Bias in Vision-Language Models with Face-Only Counterfactuals from Real Photos

DGX agent

arXiv:2601.06931v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) are increasingly deployed in socially consequential settings, raising concerns about social bias driven by demog

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

MNAFT: modality neuron-aware fine-tuning of multimodal large language models for image translation

DGX agent

arXiv:2604.16943v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have shown impressive capabilities, yet they often struggle to effectively capture the fine-grained textual inf

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

More Than Meets the Eye: Measuring the Semiotic Gap in Vision-Language Models via Semantic Anchorage

DGX agent

arXiv:2604.17354v1 Announce Type: new Abstract: Vision-Language Models (VLMs) excel at photorealistic generation, yet often struggle to represent abstract meaning such as idiomatic interpretations of

model-releasesarxiv-cs-cl
21 Apr 2026
Agents

OneDrive: Unified Multi-Paradigm Driving with Vision-Language-Action Models

DGX agent

arXiv:2604.17915v1 Announce Type: new Abstract: Vision-Language Models(VLMs) excel at autoregressive text generation, yet end-to-end autonomous driving requires multi-task learning with structured out

agentsarxiv-cs-cv
21 Apr 2026
Model Releases

Prompting Foundation Models for Zero-Shot Ship Instance Segmentation in SAR Imagery

DGX agent

arXiv:2604.17920v1 Announce Type: new Abstract: Synthetic Aperture Radar (SAR) plays a critical role in maritime surveillance, yet deep learning for SAR analysis is limited by the lack of pixel-level

model-releasesarxiv-cs-cv
21 Apr 2026
Tutorials

RePrompT: Recurrent Prompt Tuning for Integrating Structured EHR Encoders with Large Language Models

DGX agent

arXiv:2604.17725v1 Announce Type: new Abstract: Large Language Models (LLMs) have shown strong promise for mining Electronic Health Records (EHRs) by reasoning over longitudinal clinical information t

tutorialsarxiv-cs-cl
21 Apr 2026
Model Releases

SafeAnchor: Preventing Cumulative Safety Erosion in Continual Domain Adaptation of Large Language Models

DGX agent

arXiv:2604.17691v1 Announce Type: new Abstract: Safety alignment in large language models is remarkably shallow: it is concentrated in the first few output tokens and reversible by fine-tuning on as f

model-releasesarxiv-cs-lg
21 Apr 2026
Applications

SFTMix: Elevating Language Model Instruction Tuning with Mixup Recipe

DGX agent

arXiv:2410.05248v4 Announce Type: replace Abstract: To acquire instruction-following capabilities, large language models (LLMs) undergo instruction tuning, where they are trained on instruction-respon

applicationsarxiv-cs-cl
21 Apr 2026
Research

Stable Language Guidance for Vision-Language-Action Models

DGX agent

arXiv:2601.04052v2 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models have demonstrated impressive capabilities in generalized robotic control; however, they remain notoriously

researcharxiv-cs-cl
21 Apr 2026
Local Ai

Topology-Aware Layer Pruning for Large Vision-Language Models

DGX agent

arXiv:2604.16502v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated strong capabilities in natural language understanding and reasoning, while recent extensions that incorpo

local-aiarxiv-cs-cv
21 Apr 2026
Safety

Toward Consistent World Models with Multi-Token Prediction and Latent Semantic Enhancement

DGX agent

arXiv:2604.06155v2 Announce Type: replace-cross Abstract: Whether Large Language Models (LLMs) develop coherent internal world models remains a core debate. While conventional Next-Token Prediction (N

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

UniMamba: A Unified Spatial-Temporal Modeling Framework with State-Space and Attention Integration

DGX agent

arXiv:2604.16325v1 Announce Type: new Abstract: Multivariate time series forecasting is fundamental to numerous domains such as energy, finance, and environmental monitoring, where complex temporal de

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Unleashing Spatial Reasoning in Multimodal Large Language Models via Textual Representation Guided Reasoning

DGX agent

arXiv:2603.23404v2 Announce Type: replace-cross Abstract: Existing Multimodal Large Language Models (MLLMs) struggle with 3D spatial reasoning, as they fail to construct structured abstractions of the

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Unmasking the Illusion of Embodied Reasoning in Vision-Language-Action Models

DGX agent

arXiv:2604.18000v1 Announce Type: new Abstract: Recent Vision-Language-Action (VLA) models report impressive success rates on standard robotic benchmarks, fueling optimism about general-purpose physic

model-releasesarxiv-cs-ro
21 Apr 2026
Research

Aligning What Vision-Language Models See and Perceive with Adaptive Information Flow

DGX agent

arXiv:2604.15809v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have demonstrated strong capability in a wide range of tasks such as visual recognition, document parsing, and visual grou

researcharxiv-cs-cv
20 Apr 2026
Model Releases

DALM: A Domain-Algebraic Language Model via Three-Phase Structured Generation

DGX agent

arXiv:2604.15593v1 Announce Type: cross Abstract: Large language models compress heterogeneous knowledge into a single parameter space, allowing facts from different domains to interfere during genera

model-releasesarxiv-cs-ai
20 Apr 2026
Research

Learning Uncertainty from Sequential Internal Dispersion in Large Language Models

DGX agent

arXiv:2604.15741v1 Announce Type: cross Abstract: Uncertainty estimation is a promising approach to detect hallucinations in large language models (LLMs). Recent approaches commonly depend on model in

researcharxiv-cs-ai
20 Apr 2026
Research

Noise Aggregation Analysis Driven by Small-Noise Injection: Efficient Membership Inference for Diffusion Models

DGX agent

arXiv:2510.21783v2 Announce Type: replace-cross Abstract: Diffusion models have demonstrated powerful performance in generating high-quality images. A typical example is text-to-image generator like S

researcharxiv-cs-ai
20 Apr 2026
Model Releases

P3T: Prototypical Point-level Prompt Tuning with Enhanced Generalization for 3D Vision-Language Models

DGX agent

arXiv:2604.15703v1 Announce Type: new Abstract: With the rise of pre-trained models in the 3D point cloud domain for a wide range of real-world applications, adapting them to downstream tasks has beco

model-releasesarxiv-cs-cv
20 Apr 2026
Research

Protecting Language Models Against Unauthorized Distillation through Trace Rewriting

DGX agent

arXiv:2602.15143v2 Announce Type: replace Abstract: Knowledge distillation is a widely adopted technique for transferring capabilities from LLMs to smaller, more efficient student models. However, una

researcharxiv-cs-ai
20 Apr 2026
Safety

Temporal Contrastive Decoding: A Training-Free Method for Large Audio-Language Models

DGX agent

arXiv:2604.15383v1 Announce Type: cross Abstract: Large audio-language models (LALMs) generalize across speech, sound, and music, but unified decoders can exhibit a temporal smoothing bias: transient

safetyarxiv-cs-ai
20 Apr 2026
Model Releases

TRIDENT: Enhancing Large Language Model Safety with Tri-Dimensional Diversified Red-Teaming Data Synthesis

DGX agent

arXiv:2505.24672v2 Announce Type: replace Abstract: Large Language Models (LLMs) excel in various natural language processing tasks but remain vulnerable to generating harmful content or being exploit

model-releasesarxiv-cs-cl
20 Apr 2026
Model Releases

VLegal-Bench: Cognitively Grounded Benchmark for Vietnamese Legal Reasoning of Large Language Models

DGX agent

arXiv:2512.14554v5 Announce Type: replace-cross Abstract: The rapid advancement of large language models (LLMs) has enabled new possibilities for applying artificial intelligence within the legal doma

model-releasesarxiv-cs-ai
20 Apr 2026
Safety

When Search Goes Wrong: Red-Teaming Web-Augmented Large Language Models

DGX agent

arXiv:2510.09689v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have been augmented with web search to overcome the limitations of the static knowledge boundary by accessing up-

safetyarxiv-cs-ai
20 Apr 2026
Local Ai

Atropos: Improving Cost-Benefit Trade-off of LLM-based Agents under Self-Consistency with Early Termination and Model Hotswap

DGX agent

arXiv:2604.15075v1 Announce Type: cross Abstract: Open-weight Small Language Models(SLMs) can provide faster local inference at lower financial cost, but may not achieve the same performance level as

local-aiarxiv-cs-lg
17 Apr 2026
Research

CI-CBM: Class-Incremental Concept Bottleneck Model for Interpretable Continual Learning

DGX agent

arXiv:2604.14519v1 Announce Type: cross Abstract: Catastrophic forgetting remains a fundamental challenge in continual learning, in which models often forget previous knowledge when fine-tuned on a ne

researcharxiv-cs-cv
17 Apr 2026
Local Ai

Dissecting Failure Dynamics in Large Language Model Reasoning

DGX agent

arXiv:2604.14528v1 Announce Type: cross Abstract: Large Language Models (LLMs) achieve strong performance through extended inference-time deliberation, yet how their reasoning failures arise remains p

local-aiarxiv-cs-cl
17 Apr 2026
Research

DLink: Distilling Layer-wise and Dominant Knowledge from EEG Foundation Models

DGX agent

arXiv:2604.15016v1 Announce Type: new Abstract: EEG foundation models (FMs) achieve strong cross-subject and cross-task generalization but impose substantial computational and memory costs that hinder

researcharxiv-cs-lg
17 Apr 2026
← Previous
1…9394959697…1030
Next →