AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,574 results
Safety

PG-MAP: Joint MAP Optimization for Inference-Time Alignment of Diffusion and Flow-Matching Models

DGX agent

arXiv:2606.22958v1 Announce Type: cross Abstract: Inference-time alignment of pretrained text-to-image models is typically performed along a single control axis, such as classifier-free guidance, atte

safetyarxiv-cs-cv
23 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Pose Anything Anywhere:Model-free Object Poses from Arbitrary References

DGX agent

arXiv:2606.23634v1 Announce Type: new Abstract: Estimating the 6D pose of unseen objects is a fundamental yet challenging problem for open-world robotics and embodied perception. Model-based methods a

safetyarxiv-cs-cv
23 Jun 2026
Model Releases

SamatNext v0.2-B: An Exploratory Study of RMS-Normalized Hybrid Decoders for Curriculum Retention in Small Code Models

DGX agent

arXiv:2606.22248v1 Announce Type: new Abstract: Standard autoregressive Transformer decoders can often exhibit substantial forgetting under sequential fine-tuning on shifting curriculum distributions.

model-releasesarxiv-cs-lg
23 Jun 2026
Research

Skeleton-to-Image Encoding: Enabling Skeleton Representation Learning via Vision-Pretrained Models

DGX agent

arXiv:2603.05963v2 Announce Type: replace Abstract: Recent advances in large-scale pretrained vision models have demonstrated impressive capabilities across a wide range of downstream tasks, including

researcharxiv-cs-cv
23 Jun 2026
Safety

Stationary Robust Mean-Field Games under Model Mismatches

DGX agent

arXiv:2606.22579v1 Announce Type: new Abstract: Deploying multi-agent reinforcement learning (MARL) in the real world is often limited by model mismatches between the training simulators and the true

safetyarxiv-cs-lg
23 Jun 2026
Model Releases

Steer, Don't Solve: Training Small Critic Models for Large Code Agents

DGX agent

arXiv:2606.21811v1 Announce Type: cross Abstract: End-to-end code agent training is resource-intensive and plateaus on the strategy-level reasoning needed to resolve code issues, since jointly optimiz

model-releasesarxiv-cs-lg
23 Jun 2026
Applications

TaLK: Text-attributed Graph Dataset Distillation via Coupling Language Model with Graph-Aware Kernel

DGX agent

arXiv:2606.22975v1 Announce Type: new Abstract: Text-attributed graphs (TAGs) are widely used in many real-world domains, and learning on TAGs requires jointly modeling text semantics and graph struct

applicationsarxiv-cs-lg
23 Jun 2026
Research

Understanding Latent Flow Models for Tabular Data Synthesis: Targets, Paths, and Sampling

DGX agent

arXiv:2606.20878v1 Announce Type: new Abstract: Synthetic tabular data enables microdata sharing in regulated domains, yet deploying continuous-time generative models requires balancing analytical uti

researcharxiv-cs-lg
23 Jun 2026
Research

UniFS: Unified Fast-to-Slow Hierarchical Architecture for Vision-Language-Action Models

DGX agent

arXiv:2606.22794v1 Announce Type: new Abstract: Mainstream Fast-Slow dual system vision-language-action models decouple a high-frequency action expert from a low-frequency vision-language model for ef

researcharxiv-cs-ro
23 Jun 2026
Research

Unlocking In-Context Learning in Audio-Language Models from Decentralized Medical Audio

DGX agent

arXiv:2606.23243v1 Announce Type: new Abstract: Clinical audio diagnosis in low-resource settings requires models that identify conditions from minimal examples without large annotated corpora. We pro

researcharxiv-cs-lg
23 Jun 2026
Research

VeriBound: PAC-Bayesian Generalization Bounds for Process Reward Models Trained with Formal Verification Tools

DGX agent

arXiv:2606.20740v1 Announce Type: cross Abstract: Process Reward Models (PRMs) provide step-level verification for Large Language Model (LLM) reasoning, yet their training data acquisition remains a b

researcharxiv-cs-lg
23 Jun 2026
Tutorials

Words as Difference Makers: How Large Language Models Determine Causal Structure in Text

DGX agent

arXiv:2606.22430v1 Announce Type: cross Abstract: Because large language models (LLMs) are impressively successful in predicting text, it appears that they must have access to a 'world model' represen

tutorialsarxiv-cs-lg
23 Jun 2026
Tools

You can try it out on your own images here (~1.3GB model download in your browser) https://simonw.github.io/moebius-web/

DGX agent

This post highlights an interactive web-based demo of the Moebius model, which users can test directly in their browser with approximately 1.3GB of model data downloaded locally. The tool appears to b

toolssimon-willison--x
23 Jun 2026
Industry

Pretty remarkable what’s happening with open weights AI right now. We’re seeing models achieve SOTA results on specific tasks, and getting c…

DGX agent

Pretty remarkable what’s happening with open weights AI right now. We’re seeing models achieve SOTA results on specific tasks, and getting close to frontier on some areas of coding and other domains.

industryclem-delangue--x
20 Jun 2026
Research

A Judge-Aware Ranking Framework for Evaluating Large Language Models without Ground Truth

DGX agent

arXiv:2601.21817v3 Announce Type: replace-cross Abstract: Evaluating large language models (LLMs) on open-ended tasks without ground-truth labels is increasingly done via the LLM-as-a-judge paradigm.

researcharxiv-cs-lg
11 Jun 2026
Research

Afrispeech Semantics: Evaluating Audio Semantic Reasoning in Spoken Language Models Across Domains and Accents

DGX agent

arXiv:2606.11219v1 Announce Type: cross Abstract: Audio language models (ALMs) are increasingly used for speech-based understanding, yet their ability to perform semantic reasoning beyond transcriptio

researcharxiv-cs-ai
11 Jun 2026
Research

Beyond Fully Random Masking: Attention-Guided Denoising and Optimization for Diffusion Language Models

DGX agent

arXiv:2606.12273v1 Announce Type: new Abstract: Diffusion large language models (dLLMs) offer an efficient alternative to autoregressive models through parallel decoding, yet existing post-training me

researcharxiv-cs-cl
11 Jun 2026
Research

Cross-Layer Discrete Concept Discovery for Interpreting Language Models

DGX agent

arXiv:2506.20040v3 Announce Type: replace-cross Abstract: Interpreting language models remains challenging due to the existence of residual stream, which linearly mixes and duplicates features across

researcharxiv-cs-ai
11 Jun 2026
Model Releases

Damage-TriageFormer: A Foundation-Model Framework for Typology-Based Building Damage Assessment from Mono-Temporal Imagery

DGX agent

arXiv:2606.12248v1 Announce Type: new Abstract: Decision-relevant building damage assessment is critical for prioritizing resources and recovery after a disaster, yet most automated methods either fla

model-releasesarxiv-cs-cv
11 Jun 2026
Local Ai

Frozen Foundation-Model Embeddings Discard Small-Lesion Signal in Chest Radiography: Implications for Pre-Deployment Evaluation

DGX agent

arXiv:2606.11606v1 Announce Type: new Abstract: Frozen vision-transformer (ViT) foundation-model embeddings increasingly serve as the substrate for downstream chest-radiography (CXR) pipelines, yet wh

local-aiarxiv-cs-cv
11 Jun 2026
Research

Lius: Translation Model Based Instructional Lingustic Using Continual Instruction Tuning In Kupang Malay

DGX agent

arXiv:2606.11786v1 Announce Type: new Abstract: Large Language Models (LLMs) offer new potential for translation tasks but often experience performance degradation when handling low-resource languages

researcharxiv-cs-cl
11 Jun 2026
Model Releases

OmniLoc: A Geometry-Aware Foundation Model for Anchor-Free UE Localization Across Diverse Indoor Environments

DGX agent

arXiv:2606.11490v1 Announce Type: new Abstract: Indoor localization from wireless measurements remains challenging in large-scale deployments due to substantial variation in building geometry, the set

model-releasesarxiv-cs-lg
11 Jun 2026
Agents

Physics-informed generative AI for semiconductor manufacturing: Enforcing hard physical constraints in generative models by construction

DGX agent

arXiv:2606.11247v1 Announce Type: cross Abstract: Generative models are increasingly used to propose designs, data, and control actions for physical systems, yet many such systems are governed by hard

agentsarxiv-cs-ai
11 Jun 2026
Research

Pretrained self-supervised speech models can recognize unseen consonants

DGX agent

arXiv:2606.11542v1 Announce Type: cross Abstract: Modern pretrained self-supervised automatic speech recognition models are trained on large-scale audio data to encode speech into contextualized repre

researcharxiv-cs-ai
11 Jun 2026
Agents

PRInTS: Reward Modeling for Long-Horizon Information Seeking

DGX agent

arXiv:2511.19314v2 Announce Type: replace Abstract: Information-seeking is a core capability for AI agents, requiring them to gather and reason over tool-generated information across long trajectories

agentsarxiv-cs-ai
11 Jun 2026
Research

Re-evaluating Confidence Remasking in Masked Diffusion Language Models

DGX agent

arXiv:2606.12232v1 Announce Type: new Abstract: Masked diffusion language models (dLLMs) have recently emerged as a competitive alternative to autoregressive language models, with the promise of faste

researcharxiv-cs-lg
11 Jun 2026
Model Releases

Time-Series Foundation Model Embeddings for Remaining Useful Life Estimation

DGX agent

arXiv:2606.11990v1 Announce Type: cross Abstract: Remaining Useful Life (RUL) prediction is essential for industrial predictive maintenance, yet many learning-based approaches rely on extensive featur

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Towards Fully Automated Exam Grading: Fairness-Aware Recognition of Handwritten Answers with Foundation Models

DGX agent

arXiv:2606.11477v1 Announce Type: cross Abstract: Correcting handwritten exams by hand is time-consuming and error-prone, particularly for large cohorts, while fully digital exams tend to force a dida

model-releasesarxiv-cs-ai
11 Jun 2026
Applications

AI Serving Platform That Adapts to Your Model

DGX agent

Databricks offers an AI serving platform designed to flexibly accommodate different machine learning models and their specific requirements. The platform likely provides features for deploying, scalin

applicationsdatabricks
10 Jun 2026
Hardware

ASTRA-sim 3.0: Next-Level Distributed Machine Learning Simulations via High-Fidelity GPU and Infrastructure Modeling

DGX agent

arXiv:2606.10440v1 Announce Type: cross Abstract: Distributed machine learning (ML) is a key paradigm for today's large-scale artificial intelligence applications. As model inference arises as an impo

hardwarearxiv-cs-lg
10 Jun 2026
Tutorials

Breaking the Curse of Dimensionality: Diffusion Models Efficiently Learn Low-Dimensional Distributions

DGX agent

arXiv:2409.02426v5 Announce Type: replace-cross Abstract: Despite their empirical success across a wide range of generative tasks, the fundamental principles underlying the ability of diffusion models

tutorialsarxiv-cs-cv
10 Jun 2026
Safety

Flow-DPPO: Divergence Proximal Policy Optimization for Flow Matching Models

DGX agent

arXiv:2606.11025v1 Announce Type: new Abstract: Recent work has demonstrated that online reinforcement learning (RL) can substantially improve the quality and alignment of flow matching models for ima

safetyarxiv-cs-lg
10 Jun 2026
Research

GRAFT: Gain-Recalibrated Adapters for Transformer-Based Neural Population Activity Modeling

DGX agent

arXiv:2606.11066v1 Announce Type: new Abstract: Neural population activity models can recover rich temporal structure from binned spikes, but their read-in and readout layers often remain tied to a fi

researcharxiv-cs-lg
10 Jun 2026
Tools

I asked 8 AI models (including Fable 5) for their world cup predictions. Going to keep an updated leaderboard based on match results to see …

DGX agent

I asked 8 AI models (including Fable 5) for their world cup predictions. Going to keep an updated leaderboard based on match results to see which AI model performed the best! Launching tomorrow, right

toolstogether-ai--x
10 Jun 2026
Safety

In addition to transparency, I now believe frontier models should face mandatory third-party testing for cyber, bio, and autonomy risks—with…

DGX agent

In addition to transparency, I now believe frontier models should face mandatory third-party testing for cyber, bio, and autonomy risks—with the power to block or revoke deployment of models that pose

safetydario-amodei--x
10 Jun 2026
Model Releases

Instruction Finetuning DeepSeek-R1-8B Model Using LoRA and NEFTune

DGX agent

arXiv:2606.10392v1 Announce Type: new Abstract: Financial named-entity recognition (NER) is essential for translating unstructured financial reports and news into structured knowledge graphs. However,

model-releasesarxiv-cs-ai
10 Jun 2026
Applications

K-Forcing: Joint Next-K-Token Decoding via Push-Forward Language Modeling

DGX agent

arXiv:2606.10820v1 Announce Type: cross Abstract: Autoregressive (AR) language modeling is the dominant paradigm for text generation, yet its sequential token-by-token decoding makes inference memory-

applicationsarxiv-cs-ai
10 Jun 2026
Safety

MIND-V: Hierarchical World Model for Long-Horizon Robotic Manipulation with RL-based Physical Alignment

DGX agent

arXiv:2512.06628v3 Announce Type: replace-cross Abstract: Scalable embodied intelligence is constrained by the scarcity of diverse, long-horizon robotic manipulation data. Existing video world models

safetyarxiv-cs-cv
10 Jun 2026
Safety

MMD Guidance: Training-Free Distribution Adaptation for Diffusion Models via Maximum Mean Discrepancy Guidance

DGX agent

arXiv:2601.08379v2 Announce Type: replace-cross Abstract: Pre-trained diffusion models have emerged as powerful generative priors for both unconditional and conditional sample generation, yet their ou

safetyarxiv-cs-ai
10 Jun 2026
Research

Model-Based Diffusion Sampling for Predictive Control in Offline Decision Making

DGX agent

arXiv:2512.08280v3 Announce Type: replace-cross Abstract: Offline decision-making via diffusion models often produces trajectories that are misaligned with system dynamics, limiting their reliability

researcharxiv-cs-ai
10 Jun 2026
Model Releases

NOVA: Symbolic Regression Discovery of Interpretable Car-Following and Lane-Change Models with Driver Heterogeneity

DGX agent

arXiv:2606.10583v1 Announce Type: cross Abstract: We present NOVA, an autonomous symbolic regression framework that identifies interpretable car-following and lane-change structures from raw trajector

model-releasesarxiv-cs-ai
10 Jun 2026
Safety

ParaBridge: Bridging Paralinguistic Perception and Dialogue Behavior in Speech Language Models

DGX agent

arXiv:2606.10581v1 Announce Type: new Abstract: Speech carries more information than just words: a child's voice, a fearful tone, or a noisy background should all lead a sufficiently competent spoken-

safetyarxiv-cs-cl
10 Jun 2026
Model Releases

SPDM: Geometry-Modulated State Space Modeling with Manifold Constraints for Time Series Forecasting

DGX agent

arXiv:2606.09917v1 Announce Type: new Abstract: Multivariate time series forecasting requires capturing the continuously evolving correlation structure among interacting variables. Existing state-spac

model-releasesarxiv-cs-lg
10 Jun 2026
Agents

They should name the next model Halo 6 then ninja gaiden 7 cover the entire xbox catalogue

DGX agent

This post suggests naming future models after Xbox game franchises, specifically proposing 'Halo 6' and 'Ninja Gaiden 7' as model names that would reference the broader Xbox catalog. The comment appea

agentsjerry-liu--x
10 Jun 2026
Model Releases

When Do Autoregressive Sequence Models Forecast Physical Wavefields? A Controlled Study on Synthetic Seismograms

DGX agent

arXiv:2606.10868v1 Announce Type: new Abstract: Long-horizon autoregressive forecasting of oscillatory physical signals, such as seismograms, gravitational-wave strain, and similar wavefields is limit

model-releasesarxiv-cs-lg
10 Jun 2026
Safety

Anthropic played the media like a fiddle. From “untold catastrophe” to “check our latest model”, in two months and a day 🙄 (cc @tomfriedman…

DGX agent

Anthropic played the media like a fiddle. From “untold catastrophe” to “check our latest model”, in two months and a day 🙄 (cc @tomfriedman) This is the scary phase of AI — a model deemed so powerful

safetygary-marcus--x
9 Jun 2026
Model Releases

Artificial Intelligence for Mathematical Reasoning: An Integrated Survey of Language Models, Neuro-symbolic Systems, and Verified Discovery

DGX agent

arXiv:2606.08728v1 Announce Type: new Abstract: Mathematical reasoning has long served as a stringent test of machine intelligence; over the past decade, it has moved from a niche problem within NLP t

model-releasesarxiv-cs-ai
9 Jun 2026
Local Ai

Beyond Point Estimates: Benchmarking Uncertainty Quantification Methods on the AION-1 Astronomical Foundation Model

DGX agent

arXiv:2606.07771v1 Announce Type: cross Abstract: Foundation models for astronomical surveys offer powerful learned representations that can be transferred to downstream regression tasks such as galax

local-aiarxiv-cs-ai
9 Jun 2026
← Previous
1…163164165166167…1262
Next →