AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,435 results
Research

A Judge-Aware Ranking Framework for Evaluating Large Language Models without Ground Truth

DGX agent

arXiv:2601.21817v3 Announce Type: replace-cross Abstract: Evaluating large language models (LLMs) on open-ended tasks without ground-truth labels is increasingly done via the LLM-as-a-judge paradigm.

researcharxiv-cs-lg
11 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Afrispeech Semantics: Evaluating Audio Semantic Reasoning in Spoken Language Models Across Domains and Accents

DGX agent

arXiv:2606.11219v1 Announce Type: cross Abstract: Audio language models (ALMs) are increasingly used for speech-based understanding, yet their ability to perform semantic reasoning beyond transcriptio

researcharxiv-cs-ai
11 Jun 2026
Research

Beyond Fully Random Masking: Attention-Guided Denoising and Optimization for Diffusion Language Models

DGX agent

arXiv:2606.12273v1 Announce Type: new Abstract: Diffusion large language models (dLLMs) offer an efficient alternative to autoregressive models through parallel decoding, yet existing post-training me

researcharxiv-cs-cl
11 Jun 2026
Research

Cross-Layer Discrete Concept Discovery for Interpreting Language Models

DGX agent

arXiv:2506.20040v3 Announce Type: replace-cross Abstract: Interpreting language models remains challenging due to the existence of residual stream, which linearly mixes and duplicates features across

researcharxiv-cs-ai
11 Jun 2026
Model Releases

Damage-TriageFormer: A Foundation-Model Framework for Typology-Based Building Damage Assessment from Mono-Temporal Imagery

DGX agent

arXiv:2606.12248v1 Announce Type: new Abstract: Decision-relevant building damage assessment is critical for prioritizing resources and recovery after a disaster, yet most automated methods either fla

model-releasesarxiv-cs-cv
11 Jun 2026
Local Ai

Frozen Foundation-Model Embeddings Discard Small-Lesion Signal in Chest Radiography: Implications for Pre-Deployment Evaluation

DGX agent

arXiv:2606.11606v1 Announce Type: new Abstract: Frozen vision-transformer (ViT) foundation-model embeddings increasingly serve as the substrate for downstream chest-radiography (CXR) pipelines, yet wh

local-aiarxiv-cs-cv
11 Jun 2026
Research

Lius: Translation Model Based Instructional Lingustic Using Continual Instruction Tuning In Kupang Malay

DGX agent

arXiv:2606.11786v1 Announce Type: new Abstract: Large Language Models (LLMs) offer new potential for translation tasks but often experience performance degradation when handling low-resource languages

researcharxiv-cs-cl
11 Jun 2026
Model Releases

OmniLoc: A Geometry-Aware Foundation Model for Anchor-Free UE Localization Across Diverse Indoor Environments

DGX agent

arXiv:2606.11490v1 Announce Type: new Abstract: Indoor localization from wireless measurements remains challenging in large-scale deployments due to substantial variation in building geometry, the set

model-releasesarxiv-cs-lg
11 Jun 2026
Agents

Physics-informed generative AI for semiconductor manufacturing: Enforcing hard physical constraints in generative models by construction

DGX agent

arXiv:2606.11247v1 Announce Type: cross Abstract: Generative models are increasingly used to propose designs, data, and control actions for physical systems, yet many such systems are governed by hard

agentsarxiv-cs-ai
11 Jun 2026
Research

Pretrained self-supervised speech models can recognize unseen consonants

DGX agent

arXiv:2606.11542v1 Announce Type: cross Abstract: Modern pretrained self-supervised automatic speech recognition models are trained on large-scale audio data to encode speech into contextualized repre

researcharxiv-cs-ai
11 Jun 2026
Agents

PRInTS: Reward Modeling for Long-Horizon Information Seeking

DGX agent

arXiv:2511.19314v2 Announce Type: replace Abstract: Information-seeking is a core capability for AI agents, requiring them to gather and reason over tool-generated information across long trajectories

agentsarxiv-cs-ai
11 Jun 2026
Research

Re-evaluating Confidence Remasking in Masked Diffusion Language Models

DGX agent

arXiv:2606.12232v1 Announce Type: new Abstract: Masked diffusion language models (dLLMs) have recently emerged as a competitive alternative to autoregressive language models, with the promise of faste

researcharxiv-cs-lg
11 Jun 2026
Model Releases

Time-Series Foundation Model Embeddings for Remaining Useful Life Estimation

DGX agent

arXiv:2606.11990v1 Announce Type: cross Abstract: Remaining Useful Life (RUL) prediction is essential for industrial predictive maintenance, yet many learning-based approaches rely on extensive featur

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Towards Fully Automated Exam Grading: Fairness-Aware Recognition of Handwritten Answers with Foundation Models

DGX agent

arXiv:2606.11477v1 Announce Type: cross Abstract: Correcting handwritten exams by hand is time-consuming and error-prone, particularly for large cohorts, while fully digital exams tend to force a dida

model-releasesarxiv-cs-ai
11 Jun 2026
Hardware

ASTRA-sim 3.0: Next-Level Distributed Machine Learning Simulations via High-Fidelity GPU and Infrastructure Modeling

DGX agent

arXiv:2606.10440v1 Announce Type: cross Abstract: Distributed machine learning (ML) is a key paradigm for today's large-scale artificial intelligence applications. As model inference arises as an impo

hardwarearxiv-cs-lg
10 Jun 2026
Tutorials

Breaking the Curse of Dimensionality: Diffusion Models Efficiently Learn Low-Dimensional Distributions

DGX agent

arXiv:2409.02426v5 Announce Type: replace-cross Abstract: Despite their empirical success across a wide range of generative tasks, the fundamental principles underlying the ability of diffusion models

tutorialsarxiv-cs-cv
10 Jun 2026
Safety

Flow-DPPO: Divergence Proximal Policy Optimization for Flow Matching Models

DGX agent

arXiv:2606.11025v1 Announce Type: new Abstract: Recent work has demonstrated that online reinforcement learning (RL) can substantially improve the quality and alignment of flow matching models for ima

safetyarxiv-cs-lg
10 Jun 2026
Research

GRAFT: Gain-Recalibrated Adapters for Transformer-Based Neural Population Activity Modeling

DGX agent

arXiv:2606.11066v1 Announce Type: new Abstract: Neural population activity models can recover rich temporal structure from binned spikes, but their read-in and readout layers often remain tied to a fi

researcharxiv-cs-lg
10 Jun 2026
Model Releases

Instruction Finetuning DeepSeek-R1-8B Model Using LoRA and NEFTune

DGX agent

arXiv:2606.10392v1 Announce Type: new Abstract: Financial named-entity recognition (NER) is essential for translating unstructured financial reports and news into structured knowledge graphs. However,

model-releasesarxiv-cs-ai
10 Jun 2026
Applications

K-Forcing: Joint Next-K-Token Decoding via Push-Forward Language Modeling

DGX agent

arXiv:2606.10820v1 Announce Type: cross Abstract: Autoregressive (AR) language modeling is the dominant paradigm for text generation, yet its sequential token-by-token decoding makes inference memory-

applicationsarxiv-cs-ai
10 Jun 2026
Safety

MIND-V: Hierarchical World Model for Long-Horizon Robotic Manipulation with RL-based Physical Alignment

DGX agent

arXiv:2512.06628v3 Announce Type: replace-cross Abstract: Scalable embodied intelligence is constrained by the scarcity of diverse, long-horizon robotic manipulation data. Existing video world models

safetyarxiv-cs-cv
10 Jun 2026
Safety

MMD Guidance: Training-Free Distribution Adaptation for Diffusion Models via Maximum Mean Discrepancy Guidance

DGX agent

arXiv:2601.08379v2 Announce Type: replace-cross Abstract: Pre-trained diffusion models have emerged as powerful generative priors for both unconditional and conditional sample generation, yet their ou

safetyarxiv-cs-ai
10 Jun 2026
Research

Model-Based Diffusion Sampling for Predictive Control in Offline Decision Making

DGX agent

arXiv:2512.08280v3 Announce Type: replace-cross Abstract: Offline decision-making via diffusion models often produces trajectories that are misaligned with system dynamics, limiting their reliability

researcharxiv-cs-ai
10 Jun 2026
Model Releases

NOVA: Symbolic Regression Discovery of Interpretable Car-Following and Lane-Change Models with Driver Heterogeneity

DGX agent

arXiv:2606.10583v1 Announce Type: cross Abstract: We present NOVA, an autonomous symbolic regression framework that identifies interpretable car-following and lane-change structures from raw trajector

model-releasesarxiv-cs-ai
10 Jun 2026
Safety

ParaBridge: Bridging Paralinguistic Perception and Dialogue Behavior in Speech Language Models

DGX agent

arXiv:2606.10581v1 Announce Type: new Abstract: Speech carries more information than just words: a child's voice, a fearful tone, or a noisy background should all lead a sufficiently competent spoken-

safetyarxiv-cs-cl
10 Jun 2026
Model Releases

SPDM: Geometry-Modulated State Space Modeling with Manifold Constraints for Time Series Forecasting

DGX agent

arXiv:2606.09917v1 Announce Type: new Abstract: Multivariate time series forecasting requires capturing the continuously evolving correlation structure among interacting variables. Existing state-spac

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

When Do Autoregressive Sequence Models Forecast Physical Wavefields? A Controlled Study on Synthetic Seismograms

DGX agent

arXiv:2606.10868v1 Announce Type: new Abstract: Long-horizon autoregressive forecasting of oscillatory physical signals, such as seismograms, gravitational-wave strain, and similar wavefields is limit

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

Artificial Intelligence for Mathematical Reasoning: An Integrated Survey of Language Models, Neuro-symbolic Systems, and Verified Discovery

DGX agent

arXiv:2606.08728v1 Announce Type: new Abstract: Mathematical reasoning has long served as a stringent test of machine intelligence; over the past decade, it has moved from a niche problem within NLP t

model-releasesarxiv-cs-ai
9 Jun 2026
Local Ai

Beyond Point Estimates: Benchmarking Uncertainty Quantification Methods on the AION-1 Astronomical Foundation Model

DGX agent

arXiv:2606.07771v1 Announce Type: cross Abstract: Foundation models for astronomical surveys offer powerful learned representations that can be transferred to downstream regression tasks such as galax

local-aiarxiv-cs-ai
9 Jun 2026
Research

BLM-SGAN: Bidirectional Language Modeling for Semantic-Spatial Text-to-Image Generation

DGX agent

arXiv:2606.08847v1 Announce Type: cross Abstract: Despite the success of image generation from text descriptions, it still faces challenges that are difficult to overcome in domains such as natural la

researcharxiv-cs-ai
9 Jun 2026
Safety

ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning

DGX agent

arXiv:2606.08088v1 Announce Type: new Abstract: Reinforcement Learning from Verifiable Rewards (RLVR) has recently become a key paradigm for improving the reasoning abilities of Large Language Models

safetyarxiv-cs-lg
9 Jun 2026
Local Ai

CT-VAM: A Cerebello-Thalamic-Inspired Vision-Action Model for Efficient Visuomotor Control

DGX agent

arXiv:2606.09572v1 Announce Type: cross Abstract: Vision-language-action models have shown strong promise for robot manipulation, yet raw language is primarily needed to specify task intent rather tha

local-aiarxiv-cs-ai
9 Jun 2026
Research

DALE-CT: Depth-Aware Foundation Models for Computed Tomography

DGX agent

arXiv:2606.07775v1 Announce Type: new Abstract: Recent breakthroughs in self-supervised learning (SSL), such as the Latent-Euclidean Joint-Embedding Predictive Architecture (LeJEPA), alongside success

researcharxiv-cs-cv
9 Jun 2026
Model Releases

DeepMine-Mamba: Mitigating Information Dilution in Mamba-Based State Space Models for Document Image Binarization

DGX agent

arXiv:2606.08781v1 Announce Type: new Abstract: Document image binarization aims to separate foreground text from degraded backgrounds while preserving thin, broken, and low-contrast strokes. Although

model-releasesarxiv-cs-cv
9 Jun 2026
Research

DN-Hypo-Pipeline: An AI-Driven Workflow for Hypothesis Generation via Large Language Models and Scientific Explanations

DGX agent

arXiv:2606.08532v1 Announce Type: new Abstract: A scientific hypothesis is the first step in research and undergoes experimental validation, yet it also reflects a deep understanding of and reasoning

researcharxiv-cs-ai
9 Jun 2026
Research

Evaluating the Representation Space of Diffusion Models via Self-Supervised Principles

DGX agent

arXiv:2606.09718v1 Announce Type: cross Abstract: Diffusion models have demonstrated remarkable generative capabilities and have also emerged as powerful self-supervised representation learners, yet t

researcharxiv-cs-cv
9 Jun 2026
Safety

Exposing Hidden Biases in Text-to-Image Models via Automated Prompt Search

DGX agent

arXiv:2512.08724v3 Announce Type: replace Abstract: Text-to-image (TTI) diffusion models have achieved remarkable visual quality, yet they have been repeatedly shown to exhibit social biases across se

safetyarxiv-cs-lg
9 Jun 2026
Research

Failure by Interference: Language Models Make Balanced Parentheses Errors When Faulty Mechanisms Overshadow Sound Ones

DGX agent

arXiv:2507.00322v2 Announce Type: replace-cross Abstract: Despite remarkable advances in coding capabilities, language models (LMs) still struggle with simple syntactic tasks such as generating balanc

researcharxiv-cs-ai
9 Jun 2026
Applications

FF-JEPA: Long-Horizon Planning in World Models with Latent Planners

DGX agent

arXiv:2606.09311v1 Announce Type: new Abstract: Joint Embedding Predictive Architectures (JEPAs) have shown promising world modeling capabilities, enabling planning in latent space by optimizing actio

applicationsarxiv-cs-ai
9 Jun 2026
Safety

GenTSE: Enhancing Target Speaker Extraction via a Coarse-to-Fine Generative Language Model

DGX agent

arXiv:2512.20978v2 Announce Type: replace-cross Abstract: Language Model (LM)-based generative modeling has emerged as a promising direction for TSE, offering potential for improved generalization and

safetyarxiv-cs-ai
9 Jun 2026
Applications

iMaC: Translating Actions into Motion and Contact Images for Embodied World Models

DGX agent

arXiv:2606.09813v1 Announce Type: cross Abstract: Embodied world models have emerged as a pivotal paradigm for visual robotic decision-making and interactive environment simulation. However, conventio

applicationsarxiv-cs-cv
9 Jun 2026
Safety

Impacts of Histories and Models on LLM Grading: A Study in Advanced Software Engineering Courses

DGX agent

arXiv:2606.08400v1 Announce Type: cross Abstract: Graduate-level research reading report assessment creates a substantial labor burden for educators. While large language models (LLMs) hold great pote

safetyarxiv-cs-ai
9 Jun 2026
Safety

omega-EVA: Envision, Verify, and Act with Latent Interactive World Models

DGX agent

arXiv:2606.09457v1 Announce Type: new Abstract: Embodied policies typically map current observations directly to actions, leaving candidate-action consequences implicit. World models provide predictiv

safetyarxiv-cs-ro
9 Jun 2026
Research

Quantum latent distributions in deep generative models

DGX agent

arXiv:2508.19857v3 Announce Type: replace Abstract: Many successful families of generative models leverage a low-dimensional latent distribution that is mapped to a data distribution. Though simple la

researcharxiv-cs-lg
9 Jun 2026
Research

QuoVLA: Quotient Space for Vision-Language-Action Models

DGX agent

arXiv:2605.24890v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) models commonly adapt pretrained Vision-Language Models (VLMs) to robot control by mapping visual observations and lang

researcharxiv-cs-cv
9 Jun 2026
Model Releases

RTL-BenchLS: A Large-Scale Benchmark for RTL Reasoning and Generation with Large Language Models

DGX agent

arXiv:2606.08976v1 Announce Type: new Abstract: LLM-based RTL generation and reasoning is a promising direction for hardware design automation. High-quality benchmarks are critical infrastructure for

model-releasesarxiv-cs-ai
9 Jun 2026
Research

Securing Self-supervised Data Curation for Foundation Models Robustness

DGX agent

arXiv:2606.09511v1 Announce Type: new Abstract: Self-supervised data curation provides a pathway to scaling and improving the generalization capabilities of machine learning models. By leveraging self

researcharxiv-cs-cv
9 Jun 2026
Model Releases

Streaming Interventions: Can Video Large Language Models Correct Mistakes as They Occur?

DGX agent

arXiv:2606.09547v1 Announce Type: new Abstract: Learning everyday skills, like cooking a dish, relies increasingly on instructional media such as online videos. This opens the door to the use of video

model-releasesarxiv-cs-cv
9 Jun 2026
← Previous
1…129130131132133…1030
Next →