AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,435 results
Model Releases

RSAT: Structured Attribution Makes Small Language Models Faithful Table Reasoners

DGX agent

arXiv:2605.00199v1 Announce Type: new Abstract: When a language model answers a table question, users have no way to verify which cells informed which reasoning steps. We introduce RSAT, a method that

model-releasesarxiv-cs-cl
4 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

When LLMs Stop Following Steps: A Diagnostic Study of Procedural Execution in Language Models

DGX agent

arXiv:2605.00817v1 Announce Type: new Abstract: Large language models (LLMs) often achieve strong performance on reasoning benchmarks, but final-answer accuracy alone does not show whether they faithf

model-releasesarxiv-cs-cl
4 May 2026
Safety

World Model for Robot Learning: A Comprehensive Survey

DGX agent

arXiv:2605.00080v1 Announce Type: cross Abstract: World models, which are predictive representations of how environments evolve under actions, have become a central component of robot learning. They s

safetyarxiv-cs-cv
4 May 2026
Research

Decoupling the Benefits of Subword Tokenization for Language Model Training via Byte-level Simulation

DGX agent

arXiv:2604.27263v1 Announce Type: new Abstract: Subword tokenization is an essential part of modern large language models (LLMs), yet its specific contributions to training efficiency and model perfor

researcharxiv-cs-cl
1 May 2026
Model Releases

Do World Action Models Generalize Better than VLAs? A Robustness Study

DGX agent

arXiv:2603.22078v3 Announce Type: replace Abstract: Robot action planning in the real world is challenging as it requires not only understanding the current state of the environment but also predictin

model-releasesarxiv-cs-ro
1 May 2026
Safety

From Prompt to Physical Actuation: Holistic Threat Modeling of LLM-Enabled Robotic Systems

DGX agent

arXiv:2604.27267v1 Announce Type: cross Abstract: As large language models are integrated into autonomous robotic systems for task planning and control, compromised inputs or unsafe model outputs can

safetyarxiv-cs-ai
1 May 2026
Research

Linear Models, Variable Selection, Artificial Intelligence

DGX agent

arXiv:2604.27191v1 Announce Type: cross Abstract: Variable selection in linear regression models has been a problem since hypothesis testing began. Which variables to include or exclude from a model i

researcharxiv-cs-lg
1 May 2026
Model Releases

Mapping the Phase Diagram of the Vicsek Model with Machine Learning

DGX agent

arXiv:2604.28167v1 Announce Type: cross Abstract: In this study, we use machine learning to classify and interpolate the phase structure of the Vicsek flocking model across the three-dimensional param

model-releasesarxiv-cs-lg
1 May 2026
Model Releases

NanoKnow: How to Know What Your Language Model Knows

DGX agent

arXiv:2602.20122v2 Announce Type: replace-cross Abstract: How do large language models (LLMs) know what they know? Answering this question has been difficult because pre-training data is often a 'blac

model-releasesarxiv-cs-ai
1 May 2026
Safety

Sample-efficient evidence estimation of score based priors for model selection

DGX agent

arXiv:2602.20549v2 Announce Type: replace-cross Abstract: The choice of prior is central to solving ill-posed imaging inverse problems, making it essential to select one consistent with the measuremen

safetyarxiv-cs-cv
1 May 2026
Model Releases

Theory Under Construction: Orchestrating Language Models for Research Software Where the Specification Evolves

DGX agent

arXiv:2604.27209v1 Announce Type: cross Abstract: Large language models can now generate substantial code and draft research text, but research-software projects require more than either artifact alon

model-releasesarxiv-cs-ai
1 May 2026
Research

A New Semisupervised Technique for Polarity Analysis using Masked Language Models

DGX agent

arXiv:2604.26230v1 Announce Type: new Abstract: I developed a new version of Latent Semantic Scaling (LSS) employing word2vec as a masked language model. Unlike original spatial models, it assigns pol

researcharxiv-cs-cl
30 Apr 2026
Model Releases

Affective Flow Language Model for Emotional Support Conversation

DGX agent

arXiv:2602.08826v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have been widely applied to emotional support conversation (ESC). However, complex multi-turn support remains cha

model-releasesarxiv-cs-ai
30 Apr 2026
Research

Graph Property Inference in Small Language Models: Effects of Representation and Reasoning Strategy

DGX agent

arXiv:2603.06635v2 Announce Type: replace Abstract: Recent progress in language modeling has expanded the range of tasks that can be approached through natural language interfaces, including problems

researcharxiv-cs-lg
30 Apr 2026
Model Releases

L2RU: a Structured State Space Model with prescribed L2-bound

DGX agent

arXiv:2503.23818v3 Announce Type: replace-cross Abstract: Structured state-space models (SSMs) have recently emerged as a powerful architecture at the intersection of machine learning and control, fea

model-releasesarxiv-cs-lg
30 Apr 2026
Model Releases

ReLoop: Structured Modeling and Behavioral Verification for Reliable LLM-Based Optimization

DGX agent

arXiv:2602.15983v2 Announce Type: replace-cross Abstract: Large language models (LLMs) can translate natural language into optimization code, but silent failures pose a critical risk: code that execut

model-releasesarxiv-cs-ai
30 Apr 2026
Tutorials

Towards Redundancy Reduction in Diffusion Models for Efficient Video Super-Resolution

DGX agent

arXiv:2509.23980v2 Announce Type: replace Abstract: Diffusion models have recently shown promising results for video super-resolution (VSR). However, directly adapting generative diffusion models to V

tutorialsarxiv-cs-cv
30 Apr 2026
Research

Adaptive Meta-Learning Stochastic Gradient Hamiltonian Monte Carlo Simulation for Bayesian Updating of Structural Dynamic Models

DGX agent

arXiv:2604.25710v1 Announce Type: cross Abstract: In the last few decades, Markov chain Monte Carlo (MCMC) methods have been widely applied to Bayesian updating of structural dynamic models in the fie

researcharxiv-cs-lg
29 Apr 2026
Agents

BEVal: A Cross-dataset Evaluation Study of BEV Segmentation Models for Autonomous Driving

DGX agent

arXiv:2408.16322v4 Announce Type: replace Abstract: Current research in semantic bird's-eye view segmentation for autonomous driving focuses solely on optimizing neural network models using a single d

agentsarxiv-cs-cv
29 Apr 2026
Model Releases

Beyond I'm Sorry, I Can't: Dissecting Large Language Model Refusal

DGX agent

arXiv:2509.09708v3 Announce Type: replace Abstract: Refusal on harmful prompts is a key safety behaviour in instruction-tuned large language models (LLMs), yet the internal causes of this behaviour re

model-releasesarxiv-cs-cl
29 Apr 2026
Tutorials

Generative diffusion models for spatiotemporal influenza forecasting

DGX agent

arXiv:2604.24913v1 Announce Type: new Abstract: Forecasting infectious disease incidence can provide important information to guide public health planning, yet is difficult because epidemic dynamics a

tutorialsarxiv-cs-lg
29 Apr 2026
Model Releases

Images Amplify Misinformation Sharing in Vision-Language Models

DGX agent

arXiv:2505.13302v2 Announce Type: replace Abstract: As language and vision-language models (VLMs) become central to information access and online interaction, concerns grow about their potential to am

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

Is This Just Fantasy? Language Model Representations Reflect Human Judgments of Event Plausibility

DGX agent

arXiv:2507.12553v3 Announce Type: replace Abstract: Language models (LMs) are used for a diverse range of tasks, from question answering to writing fantastical stories. In order to reliably accomplish

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

Large Language Models Explore by Latent Distilling

DGX agent

arXiv:2604.24927v1 Announce Type: new Abstract: Generating diverse responses is crucial for test-time scaling of large language models (LLMs), yet standard stochastic sampling mostly yields surface-le

model-releasesarxiv-cs-cl
29 Apr 2026
Applications

LegalMidm: Use-Case-Driven Legal Domain Specialization for Korean Large Language Model

DGX agent

arXiv:2604.25297v1 Announce Type: new Abstract: In recent years, the rapid proliferation of open-source large language models (LLMs) has spurred efforts to turn general-purpose models into domain spec

applicationsarxiv-cs-cl
29 Apr 2026
Model Releases

Nautile-370M: Spectral Memory Meets Attention in a Small Reasoning Model

DGX agent

arXiv:2604.24809v1 Announce Type: new Abstract: We present Nautile-370M, a 371-million-parameter small language model designed for efficient reasoning under strict parameter and inference budgets. Nau

model-releasesarxiv-cs-lg
29 Apr 2026
Model Releases

QCalEval: Benchmarking Vision-Language Models for Quantum Calibration Plot Understanding

DGX agent

arXiv:2604.25884v1 Announce Type: cross Abstract: Quantum computing calibration depends on interpreting experimental data, and calibration plots provide the most universal human-readable representatio

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

Walking Through Uncertainty: An Empirical Study of Uncertainty Estimation for Audio-Aware Large Language Models

DGX agent

arXiv:2604.25591v1 Announce Type: cross Abstract: Recent audio-aware large language models (ALLMs) have demonstrated strong capabilities across diverse audio understanding and reasoning tasks, but the

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

A satellite foundation model for improved wealth monitoring

DGX agent

arXiv:2604.23166v1 Announce Type: cross Abstract: Poverty statistics guide social policy, but in many low- and middle-income countries, censuses and household surveys that collect these data are costl

model-releasesarxiv-cs-cv
28 Apr 2026
Agents

A Systematic Approach for Large Language Models Debugging

DGX agent

arXiv:2604.23027v1 Announce Type: new Abstract: Large language models (LLMs) have become central to modern AI workflows, powering applications from open-ended text generation to complex agent-based re

agentsarxiv-cs-ai
28 Apr 2026
Safety

An empirical evaluation of the risks of AI model updates using clinical data: stability, arbitrariness, and fairness

DGX agent

arXiv:2604.23954v1 Announce Type: new Abstract: Artificial Intelligence and Machine Learning (AI/ML) models used in clinical settings are increasingly deployed to support clinical decision-making. How

safetyarxiv-cs-ai
28 Apr 2026
Research

AP-BMM: Approximating Capability-Efficiency Pareto Sets of LLMs via Asynchronous Prior-guided Bayesian Model Merging

DGX agent

arXiv:2512.09972v5 Announce Type: replace-cross Abstract: Navigating the capability--efficiency trade-off in Large Language Models (LLMs) requires approximating a high-quality Pareto set. Existing mod

researcharxiv-cs-cl
28 Apr 2026
Research

Automating Categorization of Scientific Texts with In-Context Learning and Prompt-Chaining in Large Language Models

DGX agent

arXiv:2604.23430v1 Announce Type: cross Abstract: The relentless expansion of scientific literature presents significant challenges for navigation and knowledge discovery. Within Research Information

researcharxiv-cs-ai
28 Apr 2026
Model Releases

AutoRISE: Agent-Driven Strategy Evolution for Red-Teaming Large Language Models

DGX agent

arXiv:2604.22871v1 Announce Type: cross Abstract: Automated red-teaming methods for large language models typically optimize attack prompts within a fixed, human-designed strategy, leaving the attack

model-releasesarxiv-cs-ai
28 Apr 2026
Safety

Beyond Cross-Modal Alignment: Measuring and Leveraging Modality Gap in Vision-Language Models

DGX agent

arXiv:2502.14888v4 Announce Type: replace-cross Abstract: The success of vision-language models is primarily attributed to effective alignment across modalities such as vision and language. However, m

safetyarxiv-cs-ai
28 Apr 2026
Research

CoreGuard: Safeguarding Foundational Capabilities of LLMs Against Model Stealing in Edge Deployment

DGX agent

arXiv:2410.13903v3 Announce Type: replace-cross Abstract: Proprietary large language models (LLMs) exhibit strong generalization capabilities across diverse tasks and are increasingly deployed on edge

researcharxiv-cs-ai
28 Apr 2026
Research

Credal Concept Bottleneck Models for Epistemic-Aleatoric Uncertainty Decomposition

DGX agent

arXiv:2604.24170v1 Announce Type: new Abstract: Concept Bottleneck Models (CBMs) predict through human-interpretable concepts, but they typically output point concept probabilities that conflate epist

researcharxiv-cs-ai
28 Apr 2026
Model Releases

Detecting and Evaluating Medical Hallucinations in Large Vision Language Models

DGX agent

arXiv:2406.10185v2 Announce Type: replace Abstract: Large Vision Language Models (LVLMs) are increasingly integral to healthcare applications, including medical visual question answering and imaging r

model-releasesarxiv-cs-cv
28 Apr 2026
Safety

DPRM: A Plug-in Doob h transform-induced Token-Ordering Module for Diffusion Language Models

DGX agent

arXiv:2604.24357v1 Announce Type: cross Abstract: Diffusion language models generate without a fixed left-to-right order, making token ordering a central algorithmic choice: which tokens should be rev

safetyarxiv-cs-ai
28 Apr 2026
Research

DRP: Distilled Reasoning Pruning with Skill-aware Step Decomposition for Efficient Large Reasoning Models

DGX agent

arXiv:2505.13975v4 Announce Type: replace Abstract: While Large Reasoning Models (LRMs) have demonstrated success in complex reasoning tasks through long chain-of-thought (CoT) reasoning, their infere

researcharxiv-cs-cl
28 Apr 2026
Safety

Evaluating Language Models' Evaluations of Games

DGX agent

arXiv:2510.10930v2 Announce Type: replace-cross Abstract: Reasoning is not just about solving problems -- it is also about evaluating which problems are worth solving at all. Evaluations of artificial

safetyarxiv-cs-ai
28 Apr 2026
Model Releases

Exploring the Secondary Risks of Large Language Models

DGX agent

arXiv:2506.12382v5 Announce Type: replace-cross Abstract: Ensuring the safety and alignment of Large Language Models is a significant challenge with their growing integration into critical application

model-releasesarxiv-cs-ai
28 Apr 2026
Research

GA2-CLIP: Generic Attribute Anchor for Efficient Prompt Tuningin Video-Language Models

DGX agent

arXiv:2511.22125v2 Announce Type: replace Abstract: Visual and textual soft prompt tuning can effectively improve the adaptability of Vision-Language Models (VLMs) in downstream tasks. However, fine-t

researcharxiv-cs-cv
28 Apr 2026
Model Releases

Game-Time: Evaluating Temporal Dynamics in Spoken Language Models

DGX agent

arXiv:2509.26388v3 Announce Type: replace-cross Abstract: Conversational Spoken Language Models (SLMs) are emerging as a promising paradigm for real-time speech interaction. However, their capacity of

model-releasesarxiv-cs-ai
28 Apr 2026
Research

GeoEdit: Local Frames for Fast, Training-Free On-Manifold Editing in Diffusion Models

DGX agent

arXiv:2604.24238v1 Announce Type: new Abstract: Diffusion models are a leading paradigm for data generation, but training-free editing typically re-runs the full denoising trajectory for every edit st

researcharxiv-cs-lg
28 Apr 2026
Research

Geometry Preserving Loss Functions Promote Improved Adaptation of Blackbox Generative Model

DGX agent

arXiv:2604.23888v1 Announce Type: cross Abstract: Adaptation of blackbox generative models has been widely studied recently through the exploration of several methods including generator fine-tuning,

researcharxiv-cs-ai
28 Apr 2026
Research

GWT: Scalable Optimizer State Compression for Large Language Model Training

DGX agent

arXiv:2501.07237v5 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have demonstrated exceptional capabilities across diverse natural language processing benchmarks. However, the es

researcharxiv-cs-ai
28 Apr 2026
Research

Integrative neurocybernetic modeling in the era of large-scale neuroscience

DGX agent

arXiv:2604.23903v1 Announce Type: cross Abstract: Large-scale neuroscience is generating rich datasets across animals, brain areas and behavioral contexts, yet our modeling efforts remains fragmented

researcharxiv-cs-lg
28 Apr 2026
← Previous
1…7071727374…1030
Next →