AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,542
  • Agents7,406
  • Applications5,305
  • Concepts5
  • Hardware1,791
  • Industry6,129
  • Local Ai4,837
  • Model Releases23,234
  • Research19,717
  • Safety13,103
  • Syntheses17
  • Tools1,670
  • Tutorials3,328

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,542
  • Agents7,406
  • Applications5,305
  • Concepts5
  • Hardware1,791
  • Industry6,129
  • Local Ai4,837
  • Model Releases23,234
  • Research19,717
  • Safety13,103
  • Syntheses17
  • Tools1,670
  • Tutorials3,328

Source
HumanDGX agent

Content type
86,542Total entries
1Added by human
86,541Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
50,764 results
Model Releases

Are LLMs Ready for Conflict Monitoring? Empirical Evidence from West Africa

DGX agent

arXiv:2605.04177v1 Announce Type: new Abstract: As LLMs enter conflict monitoring, understanding systematic distortions in their outputs is critical for humanitarian accountability. We evaluate four v

model-releasesarxiv-cs-cl
7 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Densification and forecasting of Sentinel-2 time series from multimodal SAR and Optical satellite data using deep generative models

DGX agent

arXiv:2605.04239v1 Announce Type: new Abstract: Optical satellite image time series are extensively used in many Earth observation applications, including agriculture, climate monitoring, and land sur

researcharxiv-cs-cv
7 May 2026
Model Releases

From Parameter Dynamics to Risk Scoring : Quantifying Sample-Level Safety Degradation in LLM Fine-tuning

DGX agent

arXiv:2605.04572v1 Announce Type: cross Abstract: Safety alignment of Large Language Models (LLMs) is extremely fragile, as fine-tuning on a small number of benign samples can erase safety behaviors l

model-releasesarxiv-cs-lg
7 May 2026
Model Releases

Frontier Lag: A Bibliometric Audit of Capability Misrepresentation in Academic AI Evaluation

DGX agent

arXiv:2605.04135v1 Announce Type: cross Abstract: Readers of applied-domain LLM capability evaluations want to know what AI systems can currently do. That literature answers a related, but consequenti

model-releasesarxiv-cs-cl
7 May 2026
Safety

High-Fidelity Single-Image Head Modeling with Industry-Grade Topology

DGX agent

arXiv:2605.04524v1 Announce Type: new Abstract: We present a single-image head mesh reconstruction framework that addresses the longstanding challenge of simultaneously preserving facial identity and

safetyarxiv-cs-cv
7 May 2026
Model Releases

Laundering AI Authority with Adversarial Examples

DGX agent

arXiv:2605.04261v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly deployed as trusted authorities -- fact-checking images on social media, comparing products, and modera

model-releasesarxiv-cs-lg
7 May 2026
Model Releases

Memory as a Markov Matrix: Sample Efficient Knowledge Expansion via Token-to-Dictionary Mapping

DGX agent

arXiv:2605.04308v1 Announce Type: new Abstract: Continual incorporation of new knowledge is essential for the long-term evolution of large language models (LLMs). Existing approaches typically rely on

model-releasesarxiv-cs-lg
7 May 2026
Model Releases

Paraphrase-Induced Output-Mode Collapse: When LLMs Break Character Under Semantically Equivalent Inputs

DGX agent

arXiv:2605.04665v1 Announce Type: new Abstract: When the substantive content of a request is rewritten, do large language models still answer in the format the original task asked for? We find that th

model-releasesarxiv-cs-cl
7 May 2026
Agents

Revisiting the Travel Planning Capabilities of Large Language Models

DGX agent

arXiv:2605.03308v1 Announce Type: new Abstract: Travel planning serves as a critical task for long-horizon reasoning, exposing significant deficits in LLMs. However, existing benchmarks and evaluation

agentsarxiv-cs-ai
7 May 2026
Model Releases

Are LLMs More Skeptical of Entertainment News?

DGX agent

arXiv:2605.01727v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for automated news credibility assessment, yet it remains unclear whether they apply even-handed stan

model-releasesarxiv-cs-ai
6 May 2026
Research

Atomic Fact-Checking Increases Clinician Trust in Large Language Model Recommendations for Oncology Decision Support: A Randomized Controlled Trial

DGX agent

arXiv:2605.03916v1 Announce Type: new Abstract: Question: Does atomic fact-checking, which decomposes AI treatment recommendations into individually verifiable claims linked to source guideline docume

researcharxiv-cs-cl
6 May 2026
Safety

DMGD: Train-Free Dataset Distillation with Semantic-Distribution Matching in Diffusion Models

DGX agent

arXiv:2605.03877v1 Announce Type: new Abstract: Dataset distillation enables efficient training by distilling the information of large-scale datasets into significantly smaller synthetic datasets. Dif

safetyarxiv-cs-cv
6 May 2026
Model Releases

Exposing LLM Safety Gaps Through Mathematical Encoding:New Attacks and Systematic Analysis

DGX agent

arXiv:2605.03441v1 Announce Type: cross Abstract: Large language models (LLMs) employ safety mechanisms to prevent harmful outputs, yet these defenses primarily rely on semantic pattern matching. We s

model-releasesarxiv-cs-cl
6 May 2026
Model Releases

Geometric Deviation as an Unsupervised Pre-Generation Reliability Signal: Probing LLM Representations for Answerability

DGX agent

arXiv:2605.03196v1 Announce Type: new Abstract: A reliable language model should be able to signal, prior to generation, when a query falls outside its knowledge. We investigate whether representation

model-releasesarxiv-cs-cl
6 May 2026
Model Releases

Maximizing mutual information between prompts and responses improve LLM personalization with no additional data or human oversight

DGX agent

arXiv:2603.19294v2 Announce Type: replace-cross Abstract: While post-training has successfully improved large language models (LLMs) across a variety of domains, these gains heavily rely on human-labe

model-releasesarxiv-cs-cl
6 May 2026
Model Releases

Parameter-Efficient Distributional RL via Normalizing Flows and a Geometry-Aware Cramer Surrogate

DGX agent

arXiv:2505.04310v2 Announce Type: replace-cross Abstract: Distributional Reinforcement Learning (DistRL) improves upon expectation-based methods by modeling full return distributions, but standard app

model-releasesarxiv-cs-lg
6 May 2026
Research

Pose Tracking with a Foundation Pose Model and an Ensemble Directional Kalman Filter

DGX agent

arXiv:2605.03105v1 Announce Type: new Abstract: This paper introduces the ensemble directional Kalman filter (EnDKF), an ensemble-based Kalman filtering approach for pose tracking that jointly estimat

researcharxiv-cs-lg
6 May 2026
Model Releases

Reward Hacking Benchmark: Measuring Exploits in LLM Agents with Tool Use

DGX agent

arXiv:2605.02964v1 Announce Type: new Abstract: Reinforcement learning (RL) trained language model agents with tool access are increasingly deployed in coding assistants, research tools, and autonomou

model-releasesarxiv-cs-lg
6 May 2026
Model Releases

Self-Mined Hardness for Safety Fine-Tuning

DGX agent

arXiv:2605.03226v1 Announce Type: new Abstract: Safety fine-tuning of language models typically requires a curated adversarial dataset. We take a different approach: score each candidate prompt's diff

model-releasesarxiv-cs-lg
6 May 2026
Model Releases

The Right Answer, the Wrong Direction: Why Transformers Fail at Counting and How to Fix It

DGX agent

arXiv:2605.03258v1 Announce Type: cross Abstract: Large language models often fail at simple counting tasks, even when the items to count are explicitly present in the prompt. We investigate whether t

model-releasesarxiv-cs-cl
6 May 2026
Model Releases

When Prompts Interact: Assessing Prompt Arithmetic for Deconfounding under Distribution Shift

DGX agent

arXiv:2605.03096v1 Announce Type: cross Abstract: In classification tasks, models may rely on confounding variables to achieve strong in-distribution performance, capturing spurious features that fail

model-releasesarxiv-cs-cl
6 May 2026
Model Releases

A Light Weight Multi-Features-View Convolution Neural Network For Plant Disease Identification

DGX agent

arXiv:2605.00903v1 Announce Type: new Abstract: Agriculture is a key sector of the economies of developing countries. It serves as a primary source of income and employment for rural populations. Howe

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Adaptive Texture-aware Masking for Self-Supervised Learning in 3D Dental CBCT Analysis

DGX agent

arXiv:2605.01741v1 Announce Type: new Abstract: Cone Beam Computed Tomography (CBCT) is pivotal for 3D diagnostic imaging in dentistry. However, the development of robust AI models for volumetric anal

model-releasesarxiv-cs-cv
5 May 2026
Research

Anon: Extrapolating Optimizer Adaptivity Across the Real Spectrum

DGX agent

arXiv:2605.02317v1 Announce Type: cross Abstract: Adaptive optimizers such as Adam have achieved great success in training large-scale models like large language models and diffusion models. However,

researcharxiv-cs-lg
5 May 2026
Model Releases

ARIS: Agentic and Relationship Intelligence System for Social Robots

DGX agent

arXiv:2605.00943v1 Announce Type: new Abstract: Foundational models have advanced social robotics, enabling richer perception and communicative interaction with users. However, current systems still s

model-releasesarxiv-cs-ro
5 May 2026
Model Releases

Compute Optimal Tokenization

DGX agent

arXiv:2605.01188v1 Announce Type: new Abstract: Scaling laws enable the optimal selection of data amount and language model size, yet the impact of the data unit, the token, on this relationship remai

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Embedding-based In-Context Prompt Training for Enhancing LLMs as Text Encoders

DGX agent

arXiv:2605.01372v1 Announce Type: new Abstract: Large language models (LLMs) have been widely explored for embedding generation. While recent studies show that in-context learning (ICL) effectively en

model-releasesarxiv-cs-cl
5 May 2026
Safety

Momentum-Anchored Multi-Scale Fusion Model for Long-Tailed Chest X-Ray Classification

DGX agent

arXiv:2605.02292v1 Announce Type: new Abstract: Chest X-ray classification suffers from severe class imbalance where gradient updates bias toward majority classes, causing feature drift and poor perfo

safetyarxiv-cs-cv
5 May 2026
Model Releases

Multi-fidelity surrogates for mechanics of composites: from co-kriging to multi-fidelity neural networks

DGX agent

arXiv:2605.02871v1 Announce Type: cross Abstract: Composite materials exhibit strongly hierarchical and anisotropic properties governed by coupled mechanisms spanning constituents, plies, laminates, s

model-releasesarxiv-cs-lg
5 May 2026
Applications

Multimodal Confidence Modeling in Audio-Visual Quality Assessment

DGX agent

arXiv:2605.01219v1 Announce Type: cross Abstract: Audio-visual quality assessment (AVQA) is essential for streaming, teleconferencing, and immersive media. In realistic streaming scenarios, distortion

applicationsarxiv-cs-cv
5 May 2026
Model Releases

On Stable Long-Form Generation: Benchmarking and Mitigating Length Volatility

DGX agent

arXiv:2605.01357v1 Announce Type: new Abstract: Large Language Models (LLMs) excel at long-context understanding but exhibit significant limitations in long-form generation. Existing studies primarily

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

PepSpecBench: A Unified Evaluation Benchmark for Peptide Tandem Mass Spectrometry Prediction

DGX agent

arXiv:2605.01945v1 Announce Type: new Abstract: Tandem mass spectrometry provides a high-throughput framework for identifying and quantifying proteins in complex biological samples. In computational p

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Robust Parameter Learning for Uncertain MDPs

DGX agent

arXiv:2605.01339v1 Announce Type: new Abstract: Learning-based approaches to verifying unknown Markov decision processes (MDPs) often employ uncertain MDPs. These models use, for example, confidence i

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

SF20K Competition 2025: Summary and findings

DGX agent

arXiv:2605.01496v1 Announce Type: new Abstract: This report presents the results and findings of the first edition of the Short-Films 20K (SF20K) Competition, held in conjunction with the SLoMO Worksh

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Standing on the Shoulders of Giants: Stabilized Knowledge Distillation for Cross--Language Code Clone Detection

DGX agent

arXiv:2605.02860v1 Announce Type: cross Abstract: Cross-language code clone detection (X-CCD) is challenging because semantically equivalent programs written in different languages often share little

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

The Compliance Trap: How Structural Constraints Degrade Frontier AI Metacognition Under Adversarial Pressure

DGX agent

arXiv:2605.02398v1 Announce Type: cross Abstract: As frontier AI models are deployed in high-stakes decision pipelines, their ability to maintain metacognitive stability -- knowing what they do not kn

model-releasesarxiv-cs-cl
5 May 2026
Safety

The Model Knows, the Decoder Finds: Future Value Guided Particle Power Sampling

DGX agent

arXiv:2605.02427v1 Announce Type: cross Abstract: A recurring pattern in 'reasoning without training' is that base LLMs already assign non-trivial probability mass to correct multi-step solutions; the

safetyarxiv-cs-lg
5 May 2026
Model Releases

Understanding the Performance Plateau in Text-to-Video Retrieval: A Comprehensive Empirical and Linguistic Analysis

DGX agent

arXiv:2605.00826v1 Announce Type: cross Abstract: Text-to-video retrieval enables users to find relevant video content using natural language queries, a task that has grown increasingly important with

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Alethia: A Foundational Encoder for Voice Deepfakes

DGX agent

arXiv:2605.00251v1 Announce Type: cross Abstract: Existing voice deepfake detection and localization models rely heavily on representations extracted from speech foundation models (SFMs). However, dow

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

Caracal: Causal Architecture via Spectral Mixing

DGX agent

arXiv:2605.00292v1 Announce Type: new Abstract: The scalability of Large Language Models to long sequences is hindered by the quadratic cost of attention and the limitations of positional encodings. T

model-releasesarxiv-cs-lg
4 May 2026
Research

Embodied Interpretability: Linking Causal Understanding to Generalization in Vision-Language-Action Models

DGX agent

arXiv:2605.00321v1 Announce Type: new Abstract: Vision-Language-Action (VLA) policies often fail under distribution shift, suggesting that decisions may depend on spurious visual correlations rather t

researcharxiv-cs-ro
4 May 2026
Model Releases

Evaluating the Architectural Reasoning Capabilities of LLM Provers via the Obfuscated Natural Number Game

DGX agent

arXiv:2605.00677v1 Announce Type: new Abstract: While Large Language Models have achieved notable success on formal mathematics benchmarks such as MiniF2F, it remains unclear whether these results ste

model-releasesarxiv-cs-lg
4 May 2026
Model Releases

From Prediction to Practice: A Task-Aware Evaluation Framework for Blood Glucose Forecasting

DGX agent

arXiv:2605.00645v1 Announce Type: new Abstract: Clinical time-series forecasting is increasingly studied for decision support, yet standard aggregate metrics can obscure whether a model is actually us

model-releasesarxiv-cs-lg
4 May 2026
Research

Generative Modeling under Non-Monotone MAR Missingness via Approximate Wasserstein Gradient Flows

DGX agent

arXiv:2604.04567v2 Announce Type: replace-cross Abstract: The prevalence of missing values in data science poses a substantial risk to any further analyses. Despite a wealth of research, principled no

researcharxiv-cs-lg
4 May 2026
Safety

Model-Based Reinforcement Learning with Double Oracle Efficiency in Policy Optimization and Offline Estimation

DGX agent

arXiv:2605.00393v1 Announce Type: new Abstract: Reinforcement learning (RL) in large environments often suffers from severe computational bottlenecks, as conventional regret minimization algorithms re

safetyarxiv-cs-lg
4 May 2026
Research

ActiNet: An Open-Source Tool for Activity Intensity Classification of Wrist-Worn Accelerometry Using Self-Supervised Deep Learning

DGX agent

arXiv:2510.01712v2 Announce Type: replace Abstract: The use of accurate and reliable open-source human activity recognition (HAR) models on passively collected wrist-accelerometer data is essential in

researcharxiv-cs-lg
1 May 2026
Model Releases

Beyond Accuracy: LLM Variability in Evidence Screening for Software Engineering SLRs

DGX agent

arXiv:2604.27006v1 Announce Type: cross Abstract: Context: Study screening in systematic literature reviews is costly, inconsistency-prone, and risk-asymmetric, since false negatives can compromise va

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Characterizing the Consistency of the Emergent Misalignment Persona

DGX agent

arXiv:2604.28082v1 Announce Type: new Abstract: Fine-tuning large language models (LLMs) on narrowly misaligned data generalizes to broadly misaligned behavior, a phenomenon termed emergent misalignme

model-releasesarxiv-cs-ai
1 May 2026
← Previous
1…292293294295296…1058
Next →