AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,020
  • Agents7,759
  • Applications5,540
  • Concepts5
  • Hardware1,925
  • Industry6,204
  • Local Ai5,102
  • Model Releases24,783
  • Research20,783
  • Safety13,742
  • Syntheses17
  • Tools1,680
  • Tutorials3,480

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,020
  • Agents7,759
  • Applications5,540
  • Concepts5
  • Hardware1,925
  • Industry6,204
  • Local Ai5,102
  • Model Releases24,783
  • Research20,783
  • Safety13,742
  • Syntheses17
  • Tools1,680
  • Tutorials3,480

Source
HumanDGX agent

Content type
91,020Total entries
1Added by human
91,019Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,766 results
Model Releases

EgoCoT-Bench: Benchmarking Grounded and Verifiable Operation-Centric Chain of Thought Reasoning for MLLMs

DGX agent

arXiv:2605.19559v1 Announce Type: cross Abstract: The rapid development of Multimodal Large Language Models (MLLMs) has led to growing interest in egocentric video understanding, specifically the abil

model-releasesarxiv-cs-ai
20 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Fingerprinting LLMs via Prompt Injection

DGX agent

arXiv:2509.25448v3 Announce Type: replace-cross Abstract: Large language models (LLMs) are often modified after release through post-processing such as post-training or quantization, which makes it ch

researcharxiv-cs-cl
20 May 2026
Model Releases

How Faithful Is Trajectory-Based Data Attribution? Error Sources, Remedies, and Practical Guidelines

DGX agent

arXiv:2605.18814v1 Announce Type: new Abstract: Trajectory-based data attribution methods estimate the influence of training samples on model predictions by unrolling the training trajectory. They are

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

LMM-Track4D: Eliciting 4D Dynamic Reasoning in LMMs via Trajectory-Grounded Dialogue

DGX agent

arXiv:2605.19390v1 Announce Type: new Abstract: Recent large multimodal models (LMMs) have become increasingly capable on image and video understanding, yet still struggle to sustain 4D continuous spa

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

MTraining: Distributed Dynamic Sparse Attention for Efficient Ultra-Long Context Training

DGX agent

arXiv:2510.18830v2 Announce Type: replace Abstract: The adoption of long context windows has become a standard feature in Large Language Models (LLMs), as extended contexts significantly enhance their

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

PixVerve: Advancing Native UHR Image Generation to 100MP with a Large-Scale High-Quality Dataset

DGX agent

arXiv:2605.20147v1 Announce Type: new Abstract: Text-to-Image (T2I) models have recently seen notable progress around 1K and 2K resolution. With the extreme desire for better visual experience and the

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

PrAda: Few-Shot Visual Adaptation for Text-Prompted Segmentation

DGX agent

arXiv:2605.19623v1 Announce Type: new Abstract: Segmenting images is critical for visual understanding but demands extensive pixel-level annotations. Foundational models have enabled new paradigms for

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

Provable Fairness Repair for Deep Neural Networks

DGX agent

arXiv:2605.19549v1 Announce Type: cross Abstract: Deep neural networks (DNNs) are suffering from ethical issues such as individual discrimination. In response, extensive NN repair techniques have been

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

Quantifying the Generalization Gap in Seizure Detection: A Large-Scale Empirical Benchmark via the SzCORE Challenge

DGX agent

arXiv:2505.18191v2 Announce Type: replace-cross Abstract: Reliable automatic seizure detection from long-term electroencephalography (EEG) remains an unsolved challenge, as current models often fail t

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

SAGA: A Sequence-Adaptive Generative Architecture for Multi-Horizon Probabilistic Forecasting with Adaptive Temporal Conformal Prediction

DGX agent

arXiv:2605.19014v1 Announce Type: new Abstract: Microsimulation models used by ministries of finance and central banks rely on parametric processes for lifetime earnings that capture only first and se

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

STAR: Semantic-Tuned and Tail-Adaptive Retriever for Graph-Augmented Generation

DGX agent

arXiv:2605.18765v1 Announce Type: cross Abstract: To augment Large Language Models (LLMs) for multi-hop question answering, a mainstream solution within Graph Retrieval Augmented Generation (GraphRAG)

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

ZeroSearch: Incentivize the Search Capability of LLMs without Searching

DGX agent

arXiv:2505.04588v3 Announce Type: replace Abstract: Effective information searching is essential for enhancing the reasoning and generation capabilities of large language models (LLMs). Recent researc

model-releasesarxiv-cs-cl
20 May 2026
Research

A Distributional View for Visual Mechanistic Interpretability: KL-Minimal Soft-Constraint Principle

DGX agent

arXiv:2605.17504v1 Announce Type: cross Abstract: Most current paradigms in visual mechanistic interpretability (MI) remain confined to interpreting internal units of the vision model via heuristic me

researcharxiv-cs-ai
19 May 2026
Safety

Actionable World Representation

DGX agent

arXiv:2605.18743v1 Announce Type: new Abstract: Inspired by the emergent behaviors in large language models that generalized human intelligence, the research community is pursuing similar emergent cap

safetyarxiv-cs-ai
19 May 2026
Safety

Agent Bazaar: Enabling Economic Alignment in Multi-Agent Marketplaces

DGX agent

arXiv:2605.17698v1 Announce Type: new Abstract: The deployment of Large Language Models (LLMs) as autonomous economic agents introduces systemic risks that extend beyond individual capability failures

safetyarxiv-cs-lg
19 May 2026
Model Releases

AI4BayesCode: From Natural Language Descriptions to Validated Modular Stateful Bayesian Samplers

DGX agent

arXiv:2605.18476v1 Announce Type: cross Abstract: Coding and computation remain major bottlenecks in Markov chain Monte Carlo (MCMC) workflows, especially as modern sampling algorithms have become inc

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Alignment Dynamics in LLM Fine-Tuning

DGX agent

arXiv:2605.18309v1 Announce Type: cross Abstract: Although Large Language Models (LLMs) achieve strong alignment through supervised fine-tuning and reinforcement learning from human feedback, the alig

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Benchmarking Mythos-Linked Bug Rediscovery

DGX agent

arXiv:2605.17416v1 Announce Type: cross Abstract: Anthropic's April 2026 Mythos materials combine benchmark claims with concrete bug-finding stories across OpenBSD, FreeBSD, Linux, FFmpeg, and browser

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Bridging Data Trials and Task Barriers: A Unified Framework for Sketch Biometric Identification

DGX agent

arXiv:2605.17367v1 Announce Type: new Abstract: Different from existing cross-modality identification tasks (e.g., heterogeneous face recognition, sketch re-identification, etc.), we introduce a novel

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

CLAP: Contrastive Latent-space Prompt Optimization for End-to-end Autonomous Driving

DGX agent

arXiv:2605.17284v1 Announce Type: cross Abstract: End-to-end autonomous driving systems powered by Vision-Language-Action (VLA) models achieve strong performance on common driving scenarios, yet remai

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Compounding Disadvantage: Auditing Intersectional Bias in LLM-Generated Explanations Across Indian and American STEM Education

DGX agent

arXiv:2601.14506v3 Announce Type: replace-cross Abstract: Large language models are increasingly deployed in STEM education for personalized instruction and feedback across institutions in high- and l

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

Continuous Diffusion Scales Competitively with Discrete Diffusion for Language

DGX agent

arXiv:2605.18530v1 Announce Type: cross Abstract: While diffusion has drawn considerable recent attention from the language modeling community, continuous diffusion has appeared less scalable than dis

model-releasesarxiv-cs-ai
19 May 2026
Research

Drift Flow Matching

DGX agent

arXiv:2605.17244v1 Announce Type: cross Abstract: Iterative generative models such as Flow Matching and Diffusion models have demonstrated strong test-time scaling behavior, where additional inference

researcharxiv-cs-ai
19 May 2026
Model Releases

DSAA: Dual-Stage Attribute Activation for Fine-grained Open Vocabulary Detection

DGX agent

arXiv:2605.18023v1 Announce Type: new Abstract: Open-Vocabulary Object Detection (OVD) models break the limitations of closed-set detection, enabling the iden- tification of unseen categories through

model-releasesarxiv-cs-cv
19 May 2026
Safety

Efficient Bilevel Optimization for Meta Label Correction in Noisy Label Learning

DGX agent

arXiv:2605.17833v1 Announce Type: cross Abstract: Training a deep neural network with noisy labels could reduce data annotation cost but may introduce noise into the learned model. In meta label corre

safetyarxiv-cs-ai
19 May 2026
Model Releases

Evolve the Method, Not the Prompts: Evolutionary Synthesis of Jailbreak Attacks on LLMs

DGX agent

arXiv:2511.12710v2 Announce Type: replace Abstract: Automated red teaming frameworks for Large Language Models (LLMs) have become increasingly sophisticated, yet many still formulate attack optimizati

model-releasesarxiv-cs-cl
19 May 2026
Research

Forecasting Downstream Performance of LLMs With Proxy Metrics

DGX agent

arXiv:2605.18607v1 Announce Type: new Abstract: Progress in language model development is often driven by comparative decisions: which architecture to adopt, which pretraining corpus to use, or which

researcharxiv-cs-cl
19 May 2026
Tutorials

Forward-Learned Discrete Diffusion: Learning how to noise to denoise faster

DGX agent

arXiv:2605.18204v1 Announce Type: cross Abstract: Discrete diffusion models are a powerful class of generative models with strong performance across many domains. For efficiency, however, discrete dif

tutorialsarxiv-cs-lg
19 May 2026
Model Releases

Gated KalmaNet: A Fading Memory Layer Through Test-Time Ridge Regression

DGX agent

arXiv:2511.21016v3 Announce Type: replace-cross Abstract: Linear State-Space Models (SSMs) offer an efficient alternative to softmax Attention with constant memory and linear compute, but their lossy,

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

Gemini 3.5 Flash might be fast enough for gen AI to make sense

DGX agent

Gemini 3.5 Flash runs 4x faster than other frontier models in output tokens per second while delivering frontier-level intelligence, proving that speed and quality no longer require trade-offs. The mo

model-releasesars-technica
19 May 2026
Model Releases

General Preference Reinforcement Learning

DGX agent

arXiv:2605.18721v1 Announce Type: cross Abstract: Post-training has split large language model (LLM) alignment into two largely disconnected tracks. Online reinforcement learning (RL) with verifiable

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

HPC-LLM: Practical Domain Adaptation and Retrieval-Augmented Generation for HPC Support

DGX agent

arXiv:2605.16347v1 Announce Type: new Abstract: Modern scientific research increasingly depends on High-Performance Computing (HPC) infrastructures, yet many researchers face significant operational b

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

iMiGUE-3K: A Large-Scale Benchmark for Micro-Gesture Analysis with Self-Supervised Learning

DGX agent

arXiv:2605.17179v1 Announce Type: new Abstract: Emotion understanding is a fundamental challenge in affective computing and artificial intelligence. While existing approaches predominantly focus on fa

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

Learning Faster with Better Tokens: Parameter-Efficient Vocabulary Adaptation for Specialized Text Summarization

DGX agent

arXiv:2605.17379v1 Announce Type: cross Abstract: Large language models pretrained on general-domain corpora often exhibit tokenization inefficiencies when applied to specialized domains. Although con

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

LLMs in Qualitative Research: Opportunities, Limitations, and Practical Considerations

DGX agent

arXiv:2605.16538v1 Announce Type: cross Abstract: This paper examines the opportunities, limitations, and practical considerations associated with the use of large language models (LLMs) in qualitativ

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

MANTA: Multi-turn Assessment for Nonhuman Thinking & Alignment

DGX agent

arXiv:2605.16301v1 Announce Type: cross Abstract: Single-turn benchmarks such as AnimalHarmBench (AHB) have established important baselines for measuring animal welfare alignment in large language mod

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Mixture of Experts for Low-Resource LLMs

DGX agent

arXiv:2605.17598v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) architectures enable efficient model scaling, yet expert routing behavior across underrepresented languages remains poorly unde

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

NewsLens: A Multi-Agent Framework for Adversarial News Bias Navigation

DGX agent

arXiv:2605.17364v1 Announce Type: new Abstract: Media bias detection has predominantly been framed as a classification task: assign a political label to an article or outlet. We argue this framing is

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

PESD-TSF: A Period-Aware and Explicit Structured Decomposition Framework for Long-Term Time Series Forecasting

DGX agent

arXiv:2605.16449v1 Announce Type: cross Abstract: Deep forecasting models often suffer from attenuated periodic perception and entangled trend-noise representations as network depth increases. Moreove

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

PRISMat: Policy-Driven, Permutation-Invariant Autoregressive Material Generation

DGX agent

arXiv:2605.16612v1 Announce Type: new Abstract: Rapid identification of candidate materials with target properties has become a key task in materials science. Machine learning has emerged as an altern

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

ProfBench: Multi-Domain Rubrics requiring Professional Knowledge to Answer and Judge

DGX agent

arXiv:2510.18941v2 Announce Type: replace-cross Abstract: Evaluating progress in large language models (LLMs) is often constrained by the challenge of verifying responses, limiting assessments to task

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

RAP: Runtime Adaptive Pruning for LLM Inference

DGX agent

arXiv:2505.17138v5 Announce Type: replace-cross Abstract: Large language models (LLMs) excel at language understanding and generation, but their enormous computational and memory requirements hinder d

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Scaling Laws for Code: A More Data-Hungry Regime

DGX agent

arXiv:2510.08702v2 Announce Type: replace Abstract: Code Large Language Models (LLMs) are revolutionizing software engineering. However, scaling laws that guide the efficient training are predominantl

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

SCICONVBENCH: Benchmarking LLMs on Multi-Turn Clarification for Task Formulation in Computational Science

DGX agent

arXiv:2605.18630v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed as scientific AI as- sistants, and a growing body of benchmarks evaluates their capabilities acro

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

SPATIOROUTE: Dynamic Prompt Routing for Zero-Shot Spatial Reasoning

DGX agent

arXiv:2605.18209v1 Announce Type: cross Abstract: Spatial question answering over egocentric video is a challenging task that requires Vision-Language Models (VLMs) to reason about 3D object positions

model-releasesarxiv-cs-ai
19 May 2026
Safety

SSL4RL: Revisiting Self-supervised Learning as Intrinsic Reward for Visual-Language Reasoning

DGX agent

arXiv:2510.16416v4 Announce Type: replace-cross Abstract: Vision-language models (VLMs) have shown remarkable abilities by integrating large language models with visual inputs. However, they often fai

safetyarxiv-cs-ai
19 May 2026
Model Releases

Tensor Channel Equivariant Graph Neural Networks for Molecular Polarizability Prediction

DGX agent

arXiv:2605.16891v1 Announce Type: new Abstract: We introduce a tensor-channel equivariant graph neural network for direct prediction of molecular polarizability tensors. Building on the efficient PaiN

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

Text2CAD-Bench: A Benchmark for LLM-based Text-to-Parametric CAD Generation

DGX agent

arXiv:2605.18430v1 Announce Type: new Abstract: Text-to-CAD generation aims to create parametric CAD models from natural language, enabling rapid prototyping and intuitive design workflows. However, e

model-releasesarxiv-cs-lg
19 May 2026
← Previous
1…414415416417418…1371
Next →