AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
48,975 results
Hardware

Fast AI Model Partition for Split Learning over Edge Networks

DGX agent

arXiv:2507.01041v4 Announce Type: replace-cross Abstract: Split learning (SL) is a distributed learning paradigm that can enable computation-intensive artificial intelligence (AI) applications by part

hardwarearxiv-cs-ai
15 Apr 2026
Research
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

PILOT: Planning via Internalized Latent Optimization Trajectories for Large Language Models

DGX agent

arXiv:2601.19917v2 Announce Type: replace Abstract: Strategic planning is critical for multi-step reasoning, yet compact Large Language Models (LLMs) often lack the capacity to formulate global strate

researcharxiv-cs-cl
15 Apr 2026
Research

Retrievals Can Be Detrimental: Unveiling the Backdoor Vulnerability of Retrieval-Augmented Diffusion Models

DGX agent

arXiv:2501.13340v4 Announce Type: replace Abstract: Diffusion models (DMs) have recently demonstrated remarkable generation capability. However, their training generally requires huge computational re

researcharxiv-cs-cv
15 Apr 2026
Safety

Task Alignment: A simple and effective proxy for model merging in computer vision

DGX agent

arXiv:2604.12935v1 Announce Type: new Abstract: Efficiently merging several models fine-tuned for different tasks, but stemming from the same pretrained base model, is of great practical interest. Des

safetyarxiv-cs-cv
15 Apr 2026
Model Releases

When Self-Reference Fails to Close: Matrix-Level Dynamics in Large Language Models

DGX agent

arXiv:2604.12128v1 Announce Type: new Abstract: We investigate how self-referential inputs alter the internal matrix dynamics of large language models. Measuring 106 scalar metrics across up to 7 anal

model-releasesarxiv-cs-cl
15 Apr 2026
Tutorials

A Mechanistic Analysis of Looped Reasoning Language Models

DGX agent

arXiv:2604.11791v1 Announce Type: cross Abstract: Reasoning has become a central capability in large language models. Recent research has shown that reasoning performance can be improved by looping an

tutorialsarxiv-cs-ai
14 Apr 2026
Tutorials

ActDistill: General Action-Guided Self-Derived Distillation for Efficient Vision-Language-Action Models

DGX agent

arXiv:2511.18082v3 Announce Type: replace Abstract: Recent Vision-Language-Action (VLA) models have shown impressive flexibility and generalization, yet their deployment in robotic manipulation remain

tutorialsarxiv-cs-cv
14 Apr 2026
Local Ai

Ambiguity Detection and Elimination in Automated Executable Process Modeling

DGX agent

arXiv:2604.10884v1 Announce Type: cross Abstract: Automated generation of executable Business Process Model and Notation (BPMN) models from natural-language specifications is increasingly enabled by l

local-aiarxiv-cs-ai
14 Apr 2026
Tutorials

Can Small Training Runs Reliably Guide Data Curation? Rethinking Proxy-Model Practice

DGX agent

arXiv:2512.24503v2 Announce Type: replace-cross Abstract: Data teams at frontier AI companies routinely train small proxy models to make critical decisions about pretraining data recipes for full-scal

tutorialsarxiv-cs-ai
14 Apr 2026
Model Releases

Powerful Training-Free Membership Inference Against Autoregressive Language Models

DGX agent

arXiv:2601.12104v2 Announce Type: replace-cross Abstract: Fine-tuned language models pose significant privacy risks, as they may memorize and expose sensitive information from their training data. Mem

model-releasesarxiv-cs-ai
14 Apr 2026
Research

Pseudo-Unification: Entropy Probing Reveals Divergent Information Patterns in Unified Multimodal Models

DGX agent

arXiv:2604.10949v1 Announce Type: cross Abstract: Unified multimodal models (UMMs) were designed to combine the reasoning ability of large language models (LLMs) with the generation capability of visi

researcharxiv-cs-ai
14 Apr 2026
Research

Towards Brain MRI Foundation Models for the Clinic: Findings from the FOMO25 Challenge

DGX agent

arXiv:2604.11679v1 Announce Type: new Abstract: Clinical deployment of automated brain MRI analysis faces a fundamental challenge: clinical data is heterogeneous and noisy, and high-quality labels are

researcharxiv-cs-cv
14 Apr 2026
Model Releases

VP-VLA: Visual Prompting as an Interface for Vision-Language-Action Models

DGX agent

arXiv:2603.22003v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) models typically map visual observations and linguistic instructions directly to robotic control signals. This 'black-b

model-releasesarxiv-cs-ro
14 Apr 2026
Safety

Act or Escalate? Evaluating Escalation Behavior in Automation with Language Models

DGX agent

arXiv:2604.08588v1 Announce Type: cross Abstract: Effective automation hinges on deciding when to act and when to escalate. We model this as a decision under uncertainty: an LLM forms a prediction, es

safetyarxiv-cs-ai
13 Apr 2026
Safety

Balancing User Preferences by Social Networks: A Condition-Guided Social Recommendation Model for Mitigating Popularity Bias

DGX agent

arXiv:2405.16772v2 Announce Type: replace-cross Abstract: Social recommendation models weave social interactions into their design to provide uniquely personalized recommendation results for users. Ho

safetyarxiv-cs-lg
13 Apr 2026
Research

CLIP-Inspector: Model-Level Backdoor Detection for Prompt-Tuned CLIP via OOD Trigger Inversion

DGX agent

arXiv:2604.09101v1 Announce Type: cross Abstract: Organisations with limited data and computational resources increasingly outsource model training to Machine Learning as a Service (MLaaS) providers,

researcharxiv-cs-ai
13 Apr 2026
Model Releases

Constraining Sequential Model Editing with Editing Anchor Compression

DGX agent

arXiv:2503.00035v2 Announce Type: replace-cross Abstract: Large language models (LLMs) struggle with hallucinations due to false or outdated knowledge. Given the high resource demands of retraining th

model-releasesarxiv-cs-ai
13 Apr 2026
Research

dnaHNet: A Scalable and Hierarchical Foundation Model for Genomic Sequence Learning

DGX agent

arXiv:2602.10603v3 Announce Type: replace Abstract: Genomic foundation models have the potential to decode DNA syntax, yet face a fundamental tradeoff in their input representation. Standard fixed-voc

researcharxiv-cs-lg
13 Apr 2026
Model Releases

FluidFlow: a flow-matching generative model for fluid dynamics surrogates on unstructured meshes

DGX agent

arXiv:2604.08586v1 Announce Type: cross Abstract: Computational fluid dynamics (CFD) provides high-fidelity simulations of fluid flows but remains computationally expensive for many-query applications

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

HM-Bench: A Comprehensive Benchmark for Multimodal Large Language Models in Hyperspectral Remote Sensing

DGX agent

arXiv:2604.08884v1 Announce Type: cross Abstract: While multimodal large language models (MLLMs) have made significant strides in natural image understanding, their ability to perceive and reason over

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Improving Model Performance by Adapting the KGE Metric to Account for System Non-Stationarity

DGX agent

arXiv:2604.03906v2 Announce Type: replace Abstract: Geoscientific systems tend to be characterized by pronounced temporal non-stationarity, arising from seasonal and climatic variability in hydrometeo

model-releasesarxiv-cs-lg
13 Apr 2026
Model Releases

Temperature-Dependent Performance of Prompting Strategies in Extended Reasoning Large Language Models

DGX agent

arXiv:2604.08563v1 Announce Type: cross Abstract: Extended reasoning models represent a transformative shift in Large Language Model (LLM) capabilities by enabling explicit test-time computation for c

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Asymptotic-Preserving Neural Networks for Viscoelastic Parameter Identification in Multiscale Blood Flow Modeling

DGX agent

arXiv:2604.06287v1 Announce Type: new Abstract: Mathematical models and numerical simulations offer a non-invasive way to explore cardiovascular phenomena, providing access to quantities that cannot b

model-releasesarxiv-cs-lg
10 Apr 2026
Tutorials

BADiff: Bandwidth Adaptive Diffusion Model

DGX agent

arXiv:2510.21366v3 Announce Type: replace Abstract: In this work, we propose a novel framework to enable diffusion models to adapt their generation quality based on real-time network bandwidth constra

tutorialsarxiv-cs-cv
10 Apr 2026
Model Releases

Can Vision Language Models Judge Action Quality? An Empirical Evaluation

DGX agent

arXiv:2604.08294v1 Announce Type: cross Abstract: Action Quality Assessment (AQA) has broad applications in physical therapy, sports coaching, and competitive judging. Although Vision Language Models

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Cognitive Mismatch in Multimodal Large Language Models for Discrete Symbol Understanding

DGX agent

arXiv:2603.18472v2 Announce Type: replace-cross Abstract: Multimodal large language models (MLLMs) perform strongly on natural images, yet their ability to understand discrete visual symbols remains u

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

Cross-Lingual Transfer and Parameter-Efficient Adaptation in the Turkic Language Family: A Theoretical Framework for Low-Resource Language Models

DGX agent

arXiv:2604.06202v1 Announce Type: cross Abstract: Large language models (LLMs) have transformed natural language processing, yet their capabilities remain uneven across languages. Most multilingual mo

model-releasesarxiv-cs-ai
10 Apr 2026
Tutorials

Inference-Time Scaling of Diffusion Language Models via Trajectory Refinement

DGX agent

arXiv:2507.08390v4 Announce Type: replace Abstract: Discrete diffusion models have recently emerged as strong alternatives to autoregressive language models, matching their performance through large-s

tutorialsarxiv-cs-lg
10 Apr 2026
Model Releases

Infusion: Shaping Model Behavior by Editing Training Data via Influence Functions

DGX agent

arXiv:2602.09987v5 Announce Type: replace-cross Abstract: Influence functions are commonly used to attribute model behavior to training documents. We explore the reverse: crafting training data that i

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Invisible Influences: Investigating Implicit Intersectional Biases through Persona Engineering in Large Language Models

DGX agent

arXiv:2604.06213v1 Announce Type: cross Abstract: Large Language Models (LLMs) excel at human-like language generation but often embed and amplify implicit, intersectional biases, especially under per

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

LifeAlign: Lifelong Alignment for Large Language Models with Memory-Augmented Focalized Preference Optimization

DGX agent

arXiv:2509.17183v3 Announce Type: replace-cross Abstract: Alignment plays a crucial role in Large Language Models (LLMs) in aligning with human preferences on a specific task/domain. Traditional align

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

On Integrating Resilience and Human Oversight into LLM-Assisted Modeling Workflows for Digital Twins

DGX agent

arXiv:2603.25898v2 Announce Type: replace-cross Abstract: LLM-assisted modeling holds the potential to rapidly build executable Digital Twins of complex systems from only coarse descriptions and senso

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Orion-Lite: Distilling LLM Reasoning into Efficient Vision-Only Driving Models

DGX agent

arXiv:2604.08266v1 Announce Type: new Abstract: Leveraging the general world knowledge of Large Language Models (LLMs) holds significant promise for improving the ability of autonomous driving systems

model-releasesarxiv-cs-cv
10 Apr 2026
Safety

SMPL-GPTexture: Dual-View 3D Human Texture Estimation using Text-to-Image Generation Models

DGX agent

arXiv:2504.13378v2 Announce Type: replace-cross Abstract: Generating high-quality, photorealistic textures for 3D human avatars remains a fundamental yet challenging task in computer vision and multim

safetyarxiv-cs-cv
10 Apr 2026
Research

The Persistence of Cultural Memory: Investigating Multimodal Iconicity in Diffusion Models

DGX agent

arXiv:2511.11435v3 Announce Type: replace Abstract: The ambiguity between generalization and memorization in TTI diffusion models becomes pronounced when prompts invoke culturally shared visual refere

researcharxiv-cs-cv
10 Apr 2026
Model Releases

Toward a universal foundation model for graph-structured data

DGX agent

arXiv:2604.06391v1 Announce Type: cross Abstract: Graphs are a central representation in biomedical research, capturing molecular interaction networks, gene regulatory circuits, cell--cell communicati

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

UniversalVTG: A Universal and Lightweight Foundation Model for Video Temporal Grounding

DGX agent

arXiv:2604.08522v1 Announce Type: new Abstract: Video temporal grounding (VTG) is typically tackled with dataset-specific models that transfer poorly across domains and query styles. Recent efforts to

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

VertAX: a differentiable vertex model for learning epithelial tissue mechanics

DGX agent

arXiv:2604.06896v1 Announce Type: new Abstract: Epithelial tissues dynamically reshape through local mechanical interactions among cells, a process well captured by vertex models. Yet their many tunab

model-releasesarxiv-cs-lg
10 Apr 2026
Research

Zatom-1: A Multimodal Flow Foundation Model for 3D Molecules and Materials

DGX agent

arXiv:2602.22251v3 Announce Type: replace-cross Abstract: General-purpose 3D chemical modeling encompasses molecules and materials, requiring both generative and predictive capabilities. However, most

researcharxiv-cs-ai
10 Apr 2026
Model Releases

The Scaffold Effect in Coding Agents: Harness Choice as a Hidden Variable in Coding-Agent Evaluation

DGX agent

arXiv:2607.22585v1 Announce Type: new Abstract: Public leaderboards for coding agents typically rank systems by model name and pass rate, while the surrounding harness (the scaffold that issues tools,

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

VanillaBench: The Hidden Accuracy Cost of Adversarial Robustness

DGX agent

arXiv:2607.12545v1 Announce Type: cross Abstract: Adversarial robustness research has produced hundreds of defended models over the past decade, yet the literature almost universally reports robustnes

model-releasesarxiv-cs-cv
15 Jul 2026
Model Releases

Soft-Prompt Tuning for Fair and Efficient LLM Benchmark Evaluation

DGX agent

arXiv:2606.12117v1 Announce Type: cross Abstract: Benchmark scores often misrepresent a large language model's (LLM's) knowledge, because they rely, e.g., on the model's ability to follow specific for

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

CoDistill-GRPO: A Co-Distillation Recipe for Efficient Group Relative Policy Optimization

DGX agent

arXiv:2605.08873v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO) has emerged as a powerful algorithm for improving the reasoning capabilities of language models, but often fai

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Benchmarking System Dynamics AI Assistants: Cloud Versus Local LLMs on CLD Extraction and Discussion

DGX agent

arXiv:2604.18566v1 Announce Type: cross Abstract: We present a systematic evaluation of large language model families -- spanning both proprietary cloud APIs and locally-hosted open-source models -- o

model-releasesarxiv-cs-lg
21 Apr 2026
Safety

On the Shelf Life of Fine-Tuned LLM-Judges: Future-Proofing, Backward-Compatibility, and Question Generalization

DGX agent

arXiv:2509.23542v2 Announce Type: replace Abstract: The LLM-as-a-judge paradigm is widely used in both evaluating free-text model responses and reward modeling for model alignment and fine-tuning. Rec

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

Instance-Adaptive Parametrization for Amortized Variational Inference

DGX agent

arXiv:2604.06796v1 Announce Type: cross Abstract: Latent variable models, including variational autoencoders (VAE), remain a central tool in modern deep generative modeling due to their scalability an

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

How Can Driving World Models Do Counterfactual Prediction?

DGX agent

arXiv:2608.11601v1 Announce Type: new Abstract: Driving world models are often interpreted as counterfactual simulators for observed driving episodes: given a factual driving log, they are asked what

model-releasesarxiv-cs-cv
13 Aug 2026
Model Releases

Test-Time Hallucination Control in Large Vision-Language Models

DGX agent

arXiv:2608.11474v1 Announce Type: new Abstract: Object Hallucination in large vision-language models (LVLMs), where models generate non-factual content about input images, remains a critical barrier t

model-releasesarxiv-cs-cv
13 Aug 2026
← Previous
1…2930313233…1021
Next →