AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
48,975 results
Model Releases

Low-Rank Adaptation Redux for Large Models

DGX agent

arXiv:2604.21905v1 Announce Type: new Abstract: Low-rank adaptation (LoRA) has emerged as the de facto standard for parameter-efficient fine-tuning (PEFT) of foundation models, enabling the adaptation

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

From Data to Theory: Autonomous Large Language Model Agents for Materials Science

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2604.19789v1 Announce Type: new Abstract: We present an autonomous large language model (LLM) agent for end-to-end, data-driven materials theory development. The model can choose an equation for

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Graph-Theoretic Models for the Prediction of Molecular Measurements

DGX agent

arXiv:2604.19840v1 Announce Type: new Abstract: Graph-theoretic approaches offer simplicity, interpretability, and low computational cost for molecular property prediction. Among these, the model prop

model-releasesarxiv-cs-lg
23 Apr 2026
Model Releases

Rashomon Sets and Model Multiplicity in Federated Learning

DGX agent

arXiv:2602.09520v2 Announce Type: replace Abstract: The Rashomon set captures the collection of models that achieve near-identical empirical performance yet may differ substantially in their decision

model-releasesarxiv-cs-lg
23 Apr 2026
Model Releases

The Model Says Walk: How Surface Heuristics Override Implicit Constraints in LLM Reasoning

DGX agent

arXiv:2603.29025v2 Announce Type: replace-cross Abstract: Large language models systematically fail when a salient surface cue conflicts with an unstated feasibility constraint. We study this through

model-releasesarxiv-cs-ai
23 Apr 2026
Applications

What Language Models Know But Don't Say: Non-Generative Prior Extraction for Generalization

DGX agent

arXiv:2601.17609v2 Announce Type: replace Abstract: In domains like medicine and finance, large-scale labeled data is costly and often unavailable, leading to models trained on small datasets that str

applicationsarxiv-cs-cl
23 Apr 2026
Model Releases

Are Large Language Models Economically Viable for Industry Deployment?

DGX agent

arXiv:2604.19342v1 Announce Type: new Abstract: Generative AI-powered by Large Language Models (LLMs)-is increasingly deployed in industry across healthcare decision support, financial analytics, ente

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

Deep sprite-based image models: An analysis

DGX agent

arXiv:2604.19480v1 Announce Type: new Abstract: While foundation models drive steady progress in image segmentation and diffusion algorithms compose always more realistic images, the seemingly simple

model-releasesarxiv-cs-cv
22 Apr 2026
Model Releases

MiroThinker: Pushing the Performance Boundaries of Open-Source Research Agents via Model, Context, and Interactive Scaling

DGX agent

arXiv:2511.11793v3 Announce Type: replace Abstract: We present MiroThinker v1.0, an open-source research agent designed to advance tool-augmented reasoning and information-seeking capabilities. Unlike

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

Regression with Large Language Models for Materials and Molecular Property Prediction

DGX agent

arXiv:2409.06080v2 Announce Type: replace-cross Abstract: We demonstrate the ability of large language models (LLMs) to perform material and molecular property regression tasks, a significant deviatio

model-releasesarxiv-cs-lg
22 Apr 2026
Research

Remask, Don't Replace: Token-to-Mask Refinement in Masked Diffusion Language Models

DGX agent

arXiv:2604.18738v1 Announce Type: new Abstract: Masked diffusion language models such as LLaDA2.1 rely on Token-to-Token (T2T) editing to correct their own generation errors: whenever a different toke

researcharxiv-cs-cl
22 Apr 2026
Model Releases

VecHeart: Holistic Four-Chamber Cardiac Anatomy Modeling via Hybrid VecSets

DGX agent

arXiv:2604.19403v1 Announce Type: new Abstract: Accurate cardiac anatomy modeling requires the model to be able to handle intricate interrelations among structures. In this paper, we propose VecHeart,

model-releasesarxiv-cs-cv
22 Apr 2026
Safety

Visual Adversarial Attack on Vision-Language Models for Autonomous Driving

DGX agent

arXiv:2411.18275v2 Announce Type: replace Abstract: Vision-language models (VLMs) have significantly advanced autonomous driving (AD) by enhancing reasoning capabilities. However, these models remain

safetyarxiv-cs-cv
22 Apr 2026
Model Releases

Bi-LoRA: Efficient Sharpness-Aware Minimization for Fine-Tuning Large-Scale Models

DGX agent

arXiv:2508.19564v2 Announce Type: replace Abstract: Fine-tuning large-scale pre-trained models with limited data presents significant challenges for generalization. While Sharpness-Aware Minimization

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Counterfactual Modeling with Fine-Tuned LLMs for Health Intervention Design and Sensor Data Augmentation

DGX agent

arXiv:2601.14590v2 Announce Type: replace Abstract: Counterfactual explanations (CFEs) provide human-centric interpretability by identifying the minimal, actionable changes required to alter a machine

model-releasesarxiv-cs-lg
21 Apr 2026
Research

DeepRitzSplit Neural Operator for Phase-Field Models via Energy Splitting

DGX agent

arXiv:2604.18261v1 Announce Type: cross Abstract: The multi-scale and non-linear nature of phase-field models of solidification requires fine spatial and temporal discretization, leading to long compu

researcharxiv-cs-lg
21 Apr 2026
Model Releases

Flow marching for a generative PDE foundation model

DGX agent

arXiv:2509.18611v2 Announce Type: replace Abstract: Pretraining on large-scale collections of PDE-governed spatiotemporal trajectories has recently shown promise for building generalizable models of d

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

FOREVER: Forgetting Curve-Inspired Memory Replay for Language Model Continual Learning

DGX agent

arXiv:2601.03938v2 Announce Type: replace-cross Abstract: Continual learning (CL) for large language models (LLMs) aims to enable sequential knowledge acquisition without catastrophic forgetting. Memo

model-releasesarxiv-cs-cl
21 Apr 2026
Applications

Grokking of Diffusion Models: Case Study on Modular Addition

DGX agent

arXiv:2604.17673v1 Announce Type: new Abstract: Despite their empirical success, how diffusion models generalize remains poorly understood from a mechanistic perspective. We demonstrate that diffusion

applicationsarxiv-cs-lg
21 Apr 2026
Model Releases

IncreFA: Breaking the Static Wall of Generative Model Attribution

DGX agent

arXiv:2604.17736v1 Announce Type: new Abstract: As AI generative models evolve at unprecedented speed, image attribution has become a moving target. New diffusion, adversarial and autoregressive gener

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

On the Predictive Power of Representation Dispersion in Language Models

DGX agent

arXiv:2506.24106v2 Announce Type: replace Abstract: We show that a language model's ability to predict text is tightly linked to the breadth of its embedding space: models that spread their contextual

model-releasesarxiv-cs-cl
21 Apr 2026
Safety

One Adapts to Any: Meta Reward Modeling for Personalized LLM Alignment

DGX agent

arXiv:2601.18731v2 Announce Type: replace Abstract: Alignment of Large Language Models (LLMs) aims to align outputs with human preferences, and personalized alignment further adapts models to individu

safetyarxiv-cs-cl
21 Apr 2026
Research

SmoGVLM: A Small, Graph-enhanced Vision-Language Model

DGX agent

arXiv:2604.16517v1 Announce Type: cross Abstract: Large vision-language models (VLMs) achieve strong performance on multimodal tasks but often suffer from hallucination and poor grounding in knowledge

researcharxiv-cs-cl
21 Apr 2026
Model Releases

SpeakerSleuth: Can Large Audio-Language Models Judge Speaker Consistency across Multi-turn Dialogues?

DGX agent

arXiv:2601.04029v2 Announce Type: replace Abstract: Large Audio-Language Models (LALMs) as judges have emerged as a prominent approach for evaluating speech generation quality, yet their ability to as

model-releasesarxiv-cs-cl
21 Apr 2026
Research

Stability-Weighted Decoding for Diffusion Language Models

DGX agent

arXiv:2604.17068v1 Announce Type: new Abstract: Diffusion large language models (dLLMs) enable parallel text generation by iteratively denoising a fully masked sequence, unmasking a subset of masked t

researcharxiv-cs-cl
21 Apr 2026
Applications

TensorHub: Rethinking AI Model Hub with Tensor-Centric Compression

DGX agent

arXiv:2604.17104v1 Announce Type: cross Abstract: Modern AI models are growing rapidly in size and redundancy, leading to significant storage and distribution challenges in model hubs. We present Tens

applicationsarxiv-cs-lg
21 Apr 2026
Research

Advancing Intelligent Sequence Modeling: Evolution, Trade-offs, and Applications of State- Space Architectures from S4 to Mamba

DGX agent

arXiv:2503.18970v3 Announce Type: replace Abstract: Structured State Space Models (SSMs) have emerged as a transformative paradigm in sequence modeling, addressing critical limitations of Recurrent Ne

researcharxiv-cs-lg
20 Apr 2026
Model Releases

Do Vision-Language Models Truly Perform Vision Reasoning? A Rigorous Study of the Modality Gap

DGX agent

arXiv:2604.16256v1 Announce Type: cross Abstract: Reasoning in vision-language models (VLMs) has recently attracted significant attention due to its broad applicability across diverse downstream tasks

model-releasesarxiv-cs-cl
20 Apr 2026
Safety

Elucidating the SNR-t Bias of Diffusion Probabilistic Models

DGX agent

arXiv:2604.16044v1 Announce Type: new Abstract: Diffusion Probabilistic Models have demonstrated remarkable performance across a wide range of generative tasks. However, we have observed that these mo

safetyarxiv-cs-cv
20 Apr 2026
Research

LLaMo: Scaling Pretrained Language Models for Unified Motion Understanding and Generation with Continuous Autoregressive Tokens

DGX agent

arXiv:2602.12370v2 Announce Type: replace Abstract: Recent progress in large models has led to significant advances in unified multimodal generation and understanding. However, the development of mode

researcharxiv-cs-cv
20 Apr 2026
Model Releases

No Universal Courtesy: A Cross-Linguistic, Multi-Model Study of Politeness Effects on LLMs Using the PLUM Corpus

DGX agent

arXiv:2604.16275v1 Announce Type: new Abstract: This paper explores the response of Large Language Models (LLMs) to user prompts with different degrees of politeness and impoliteness. The Politeness T

model-releasesarxiv-cs-cl
20 Apr 2026
Research

Reward Modeling for Scientific Writing Evaluation

DGX agent

arXiv:2601.11374v2 Announce Type: replace Abstract: Scientific writing is an expert-domain task that demands deep domain knowledge, task-specific requirements and reasoning capabilities that leverage

researcharxiv-cs-cl
20 Apr 2026
Applications

Sketching the Readout of Large Language Models for Scalable Data Attribution and Valuation

DGX agent

arXiv:2604.16197v1 Announce Type: new Abstract: Data attribution and valuation are critical for understanding data-model synergy for Large Language Models (LLMs), yet existing gradient-based methods s

applicationsarxiv-cs-lg
20 Apr 2026
Local Ai

AIPC: Agent-Based Automation for AI Model Deployment with Qualcomm AI Runtime

DGX agent

arXiv:2604.14661v1 Announce Type: cross Abstract: Edge AI model deployment is a multi-stage engineering process involving model conversion, operator compatibility handling, quantization calibration, r

local-aiarxiv-cs-lg
17 Apr 2026
Model Releases

Correcting Suppressed Log-Probabilities in Language Models with Post-Transformer Adapters

DGX agent

arXiv:2604.14174v1 Announce Type: new Abstract: Alignment-tuned language models frequently suppress factual log-probabilities on politically sensitive topics despite retaining the knowledge in their h

model-releasesarxiv-cs-cl
17 Apr 2026
Research

IMPACTX: improving model performance by appropriately constraining the training with teacher explanations

DGX agent

arXiv:2502.12222v2 Announce Type: replace Abstract: The eXplainable Artificial Intelligence (XAI) research predominantly concentrates to provide explainations about AI model decisions, especially Deep

researcharxiv-cs-lg
17 Apr 2026
Model Releases

Label-efficient underwater species classification with logistic regression on frozen foundation model embeddings

DGX agent

arXiv:2604.00313v2 Announce Type: replace Abstract: Automated species classification from underwater imagery is bottlenecked by the cost of expert annotation, and supervised models trained on one data

model-releasesarxiv-cs-cv
17 Apr 2026
Safety

Reasoning Dynamics and the Limits of Monitoring Modality Reliance in Vision-Language Models

DGX agent

arXiv:2604.14888v1 Announce Type: new Abstract: Recent advances in vision language models (VLMs) offer reasoning capabilities, yet how these unfold and integrate visual and textual information remains

safetyarxiv-cs-cl
17 Apr 2026
Research

Towards Faster Language Model Inference Using Mixture-of-Experts Flow Matching

DGX agent

arXiv:2604.15009v1 Announce Type: cross Abstract: Flow matching retains the generation quality of diffusion models while enabling substantially faster inference, making it a compelling paradigm for ge

researcharxiv-cs-lg
17 Apr 2026
Model Releases

Cracking the Code of Juxtaposition: Can AI Models Understand the Humorous Contradictions

DGX agent

arXiv:2405.19088v3 Announce Type: replace Abstract: Recent advancements in large multimodal language models have demonstrated remarkable proficiency across a wide range of tasks. Yet, these models sti

model-releasesarxiv-cs-cl
16 Apr 2026
Research

Diffusion Sequence Models for Generative In-Context Meta-Learning of Robot Dynamics

DGX agent

arXiv:2604.13366v1 Announce Type: new Abstract: Accurate modeling of robot dynamics is essential for model-based control, yet remains challenging under distributional shifts and real-time constraints.

researcharxiv-cs-lg
16 Apr 2026
Research

FAST: A Synergistic Framework of Attention and State-space Models for Spatiotemporal Traffic Prediction

DGX agent

arXiv:2604.13453v1 Announce Type: new Abstract: Traffic forecasting requires modeling complex temporal dynamics and long-range spatial dependencies over large sensor networks. Existing methods typical

researcharxiv-cs-lg
16 Apr 2026
Research

Indexing Multimodal Language Models for Large-scale Image Retrieval

DGX agent

arXiv:2604.13268v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have demonstrated strong cross-modal reasoning capabilities, yet their potential for vision-only tasks remain

researcharxiv-cs-cl
16 Apr 2026
Research

Neural Chain-of-Thought Search: Searching the Optimal Reasoning Path to Enhance Large Language Models

DGX agent

arXiv:2601.11340v2 Announce Type: replace Abstract: Chain-of-Thought reasoning has significantly enhanced the problem-solving capabilities of Large Language Models. Unfortunately, current models gener

researcharxiv-cs-cl
16 Apr 2026
Model Releases

TLoRA+: A Low-Rank Parameter-Efficient Fine-Tuning Method for Large Language Models

DGX agent

arXiv:2604.13368v1 Announce Type: new Abstract: Fine-tuning large language models (LLMs) aims to adapt pre-trained models to specific tasks using relatively small and domain-specific datasets. Among P

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Benchmarking Foundation Models with Retrieval-Augmented Generation in Olympic-Level Physics Problem Solving

DGX agent

arXiv:2510.00919v3 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) with foundation models has achieved strong performance across diverse tasks, but their capacity for exper

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Climate Model Tuning with Online Synchronization-Based Parameter Estimation

DGX agent

arXiv:2510.06180v2 Announce Type: replace-cross Abstract: In climate science, the tuning of climate models is a computationally intensive problem due to the combination of the high-dimensionality of t

model-releasesarxiv-cs-lg
15 Apr 2026
Model Releases

Don't Show Pixels, Show Cues: Unlocking Visual Tool Reasoning in Language Models via Perception Programs

DGX agent

arXiv:2604.12896v1 Announce Type: new Abstract: Multimodal language models (MLLMs) are increasingly paired with vision tools (e.g., depth, flow, correspondence) to enhance visual reasoning. However, d

model-releasesarxiv-cs-cv
15 Apr 2026
← Previous
1…2829303132…1021
Next →