AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,993
  • Agents7,449
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,136
  • Local Ai4,859
  • Model Releases23,375
  • Research19,835
  • Safety13,176
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,993
  • Agents7,449
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,136
  • Local Ai4,859
  • Model Releases23,375
  • Research19,835
  • Safety13,176
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

Content type
86,993Total entries
1Added by human
86,992Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,106 results
Model Releases

Bridging Data Trials and Task Barriers: A Unified Framework for Sketch Biometric Identification

DGX agent

arXiv:2605.17367v1 Announce Type: new Abstract: Different from existing cross-modality identification tasks (e.g., heterogeneous face recognition, sketch re-identification, etc.), we introduce a novel

model-releasesarxiv-cs-cv
19 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

CLAP: Contrastive Latent-space Prompt Optimization for End-to-end Autonomous Driving

DGX agent

arXiv:2605.17284v1 Announce Type: cross Abstract: End-to-end autonomous driving systems powered by Vision-Language-Action (VLA) models achieve strong performance on common driving scenarios, yet remai

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Compounding Disadvantage: Auditing Intersectional Bias in LLM-Generated Explanations Across Indian and American STEM Education

DGX agent

arXiv:2601.14506v3 Announce Type: replace-cross Abstract: Large language models are increasingly deployed in STEM education for personalized instruction and feedback across institutions in high- and l

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

Continuous Diffusion Scales Competitively with Discrete Diffusion for Language

DGX agent

arXiv:2605.18530v1 Announce Type: cross Abstract: While diffusion has drawn considerable recent attention from the language modeling community, continuous diffusion has appeared less scalable than dis

model-releasesarxiv-cs-ai
19 May 2026
Research

Drift Flow Matching

DGX agent

arXiv:2605.17244v1 Announce Type: cross Abstract: Iterative generative models such as Flow Matching and Diffusion models have demonstrated strong test-time scaling behavior, where additional inference

researcharxiv-cs-ai
19 May 2026
Model Releases

DSAA: Dual-Stage Attribute Activation for Fine-grained Open Vocabulary Detection

DGX agent

arXiv:2605.18023v1 Announce Type: new Abstract: Open-Vocabulary Object Detection (OVD) models break the limitations of closed-set detection, enabling the iden- tification of unseen categories through

model-releasesarxiv-cs-cv
19 May 2026
Safety

Efficient Bilevel Optimization for Meta Label Correction in Noisy Label Learning

DGX agent

arXiv:2605.17833v1 Announce Type: cross Abstract: Training a deep neural network with noisy labels could reduce data annotation cost but may introduce noise into the learned model. In meta label corre

safetyarxiv-cs-ai
19 May 2026
Model Releases

Evolve the Method, Not the Prompts: Evolutionary Synthesis of Jailbreak Attacks on LLMs

DGX agent

arXiv:2511.12710v2 Announce Type: replace Abstract: Automated red teaming frameworks for Large Language Models (LLMs) have become increasingly sophisticated, yet many still formulate attack optimizati

model-releasesarxiv-cs-cl
19 May 2026
Research

Forecasting Downstream Performance of LLMs With Proxy Metrics

DGX agent

arXiv:2605.18607v1 Announce Type: new Abstract: Progress in language model development is often driven by comparative decisions: which architecture to adopt, which pretraining corpus to use, or which

researcharxiv-cs-cl
19 May 2026
Tutorials

Forward-Learned Discrete Diffusion: Learning how to noise to denoise faster

DGX agent

arXiv:2605.18204v1 Announce Type: cross Abstract: Discrete diffusion models are a powerful class of generative models with strong performance across many domains. For efficiency, however, discrete dif

tutorialsarxiv-cs-lg
19 May 2026
Model Releases

Gated KalmaNet: A Fading Memory Layer Through Test-Time Ridge Regression

DGX agent

arXiv:2511.21016v3 Announce Type: replace-cross Abstract: Linear State-Space Models (SSMs) offer an efficient alternative to softmax Attention with constant memory and linear compute, but their lossy,

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

General Preference Reinforcement Learning

DGX agent

arXiv:2605.18721v1 Announce Type: cross Abstract: Post-training has split large language model (LLM) alignment into two largely disconnected tracks. Online reinforcement learning (RL) with verifiable

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

HPC-LLM: Practical Domain Adaptation and Retrieval-Augmented Generation for HPC Support

DGX agent

arXiv:2605.16347v1 Announce Type: new Abstract: Modern scientific research increasingly depends on High-Performance Computing (HPC) infrastructures, yet many researchers face significant operational b

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

iMiGUE-3K: A Large-Scale Benchmark for Micro-Gesture Analysis with Self-Supervised Learning

DGX agent

arXiv:2605.17179v1 Announce Type: new Abstract: Emotion understanding is a fundamental challenge in affective computing and artificial intelligence. While existing approaches predominantly focus on fa

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

Learning Faster with Better Tokens: Parameter-Efficient Vocabulary Adaptation for Specialized Text Summarization

DGX agent

arXiv:2605.17379v1 Announce Type: cross Abstract: Large language models pretrained on general-domain corpora often exhibit tokenization inefficiencies when applied to specialized domains. Although con

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

LLMs in Qualitative Research: Opportunities, Limitations, and Practical Considerations

DGX agent

arXiv:2605.16538v1 Announce Type: cross Abstract: This paper examines the opportunities, limitations, and practical considerations associated with the use of large language models (LLMs) in qualitativ

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

MANTA: Multi-turn Assessment for Nonhuman Thinking & Alignment

DGX agent

arXiv:2605.16301v1 Announce Type: cross Abstract: Single-turn benchmarks such as AnimalHarmBench (AHB) have established important baselines for measuring animal welfare alignment in large language mod

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Mixture of Experts for Low-Resource LLMs

DGX agent

arXiv:2605.17598v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) architectures enable efficient model scaling, yet expert routing behavior across underrepresented languages remains poorly unde

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

NewsLens: A Multi-Agent Framework for Adversarial News Bias Navigation

DGX agent

arXiv:2605.17364v1 Announce Type: new Abstract: Media bias detection has predominantly been framed as a classification task: assign a political label to an article or outlet. We argue this framing is

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

PESD-TSF: A Period-Aware and Explicit Structured Decomposition Framework for Long-Term Time Series Forecasting

DGX agent

arXiv:2605.16449v1 Announce Type: cross Abstract: Deep forecasting models often suffer from attenuated periodic perception and entangled trend-noise representations as network depth increases. Moreove

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

PRISMat: Policy-Driven, Permutation-Invariant Autoregressive Material Generation

DGX agent

arXiv:2605.16612v1 Announce Type: new Abstract: Rapid identification of candidate materials with target properties has become a key task in materials science. Machine learning has emerged as an altern

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

ProfBench: Multi-Domain Rubrics requiring Professional Knowledge to Answer and Judge

DGX agent

arXiv:2510.18941v2 Announce Type: replace-cross Abstract: Evaluating progress in large language models (LLMs) is often constrained by the challenge of verifying responses, limiting assessments to task

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

RAP: Runtime Adaptive Pruning for LLM Inference

DGX agent

arXiv:2505.17138v5 Announce Type: replace-cross Abstract: Large language models (LLMs) excel at language understanding and generation, but their enormous computational and memory requirements hinder d

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Scaling Laws for Code: A More Data-Hungry Regime

DGX agent

arXiv:2510.08702v2 Announce Type: replace Abstract: Code Large Language Models (LLMs) are revolutionizing software engineering. However, scaling laws that guide the efficient training are predominantl

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

SCICONVBENCH: Benchmarking LLMs on Multi-Turn Clarification for Task Formulation in Computational Science

DGX agent

arXiv:2605.18630v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed as scientific AI as- sistants, and a growing body of benchmarks evaluates their capabilities acro

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

SPATIOROUTE: Dynamic Prompt Routing for Zero-Shot Spatial Reasoning

DGX agent

arXiv:2605.18209v1 Announce Type: cross Abstract: Spatial question answering over egocentric video is a challenging task that requires Vision-Language Models (VLMs) to reason about 3D object positions

model-releasesarxiv-cs-ai
19 May 2026
Safety

SSL4RL: Revisiting Self-supervised Learning as Intrinsic Reward for Visual-Language Reasoning

DGX agent

arXiv:2510.16416v4 Announce Type: replace-cross Abstract: Vision-language models (VLMs) have shown remarkable abilities by integrating large language models with visual inputs. However, they often fai

safetyarxiv-cs-ai
19 May 2026
Model Releases

Tensor Channel Equivariant Graph Neural Networks for Molecular Polarizability Prediction

DGX agent

arXiv:2605.16891v1 Announce Type: new Abstract: We introduce a tensor-channel equivariant graph neural network for direct prediction of molecular polarizability tensors. Building on the efficient PaiN

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

Text2CAD-Bench: A Benchmark for LLM-based Text-to-Parametric CAD Generation

DGX agent

arXiv:2605.18430v1 Announce Type: new Abstract: Text-to-CAD generation aims to create parametric CAD models from natural language, enabling rapid prototyping and intuitive design workflows. However, e

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

'The Whole Is Greater Than the Sum of Its Parts': A Compatibility-Aware Multi-Teacher CoT Distillation Framework

DGX agent

arXiv:2601.13992v2 Announce Type: replace-cross Abstract: Chain-of-Thought (CoT) reasoning empowers Large Language Models (LLMs) with remarkable capabilities but typically requires prohibitive paramet

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

TPV: Parameter Perturbations Through the Lens of Test Prediction Variance

DGX agent

arXiv:2512.11089v4 Announce Type: replace-cross Abstract: We introduce test prediction variance (TPV)--the first-order sensitivity of a trained model's outputs to parameter perturbations--as a unifyin

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

UCSF-PDGM-VQA: Visual Question Answering dataset for brain tumor MRI interpretation

DGX agent

arXiv:2605.17140v1 Announce Type: cross Abstract: Brain tumor diagnosis is largely dependent on Magnetic Resonance Imaging (MRI) evaluation, which requires radiologists to synthesize thousands of imag

model-releasesarxiv-cs-ai
19 May 2026
Safety

Weak-to-Strong Elicitation via Mismatched Wrong Drafts

DGX agent

arXiv:2605.17314v1 Announce Type: cross Abstract: We consider whether off-policy experience from a smaller, weaker model can elicit capability in a stronger learner that on-policy RL fine-tuning (e.g.

safetyarxiv-cs-ai
19 May 2026
Research

When Efficiency Backfires: Cascading LLMs Trigger Cascade Failure under Adversarial Attack

DGX agent

arXiv:2605.17288v1 Announce Type: cross Abstract: Large Language Model (LLM) cascade systems are designed to balance efficiency and performance by processing queries with lightweight models while sele

researcharxiv-cs-ai
19 May 2026
Applications

XCTFormer: Leveraging Cross-Channel and Cross-Time Dependencies for Enhanced Time-Series Analysis

DGX agent

arXiv:2605.18534v1 Announce Type: new Abstract: Multivariate time-series analysis involves extracting informative representations from sequences of multiple interdependent variables, supporting tasks

applicationsarxiv-cs-lg
19 May 2026
Safety

ZeroSiam: An Efficient Asymmetry for Test-Time Entropy Optimization without Collapse

DGX agent

arXiv:2509.23183v3 Announce Type: replace Abstract: Test-time entropy minimization helps adapt a model to novel environments and incentivize its reasoning capability, unleashing the model's potential

safetyarxiv-cs-lg
19 May 2026
Model Releases

A Reproducible and Physically Feasible Dynamic Parameter Identification Framework for a Low-Cost Robot Arm

DGX agent

arXiv:2605.15949v1 Announce Type: new Abstract: This paper presents a reproducible and physically feasible dynamic parameter identification framework for CRANE-X7, a low-cost robot arm driven by modul

model-releasesarxiv-cs-ro
18 May 2026
Research

A Unified Perturbation Framework for Analyzing Leaderboard Stability and Manipulation

DGX agent

arXiv:2605.15761v1 Announce Type: new Abstract: Evaluation leaderboards such as LMArena play a central role in benchmarking large language models by aggregating pairwise human preferences into model r

researcharxiv-cs-lg
18 May 2026
Model Releases

Agentic Discovery of Neural Architectures: AIRA-Compose and AIRA-Design

DGX agent

arXiv:2605.15871v1 Announce Type: new Abstract: Toward recursive self-improvement, we investigate LLM agents autonomously designing foundation models beyond standard Transformers. We introduce a dual-

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Deep Pre-Alignment for VLMs

DGX agent

arXiv:2605.15300v1 Announce Type: new Abstract: Most Vision Language Models (VLMs) directly map outputs from ViT encoders to the LLM via a lightweight projector. While effective, recent analysis sugge

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

DimMem: Dimensional Structuring for Efficient Long-Term Agent Memory

DGX agent

arXiv:2605.15759v1 Announce Type: new Abstract: Large language model (LLM) agents require long-term memory to leverage information from past interactions. However, existing memory systems often face a

model-releasesarxiv-cs-cl
18 May 2026
Model Releases

Federated Learning of Spiking Neural Networks under Heterogeneous Temporal Resolutions

DGX agent

arXiv:2605.15355v1 Announce Type: new Abstract: Spiking neural networks (SNNs) are biologically inspired energy-efficient models that use sparse binary spike-based communication between neurons, makin

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

Flash-GRPO: Efficient Alignment for Video Diffusion via One-Step Policy Optimization

DGX agent

arXiv:2605.15980v1 Announce Type: new Abstract: Group Relative Policy Optimization has emerged as essential for aligning video diffusion models with human preferences, but faces a critical computation

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

FORGE: Self-Evolving Agent Memory With No Weight Updates via Population Broadcast

DGX agent

arXiv:2605.16233v1 Announce Type: new Abstract: Can LLM agents improve decision-making through self-generated memory without gradient updates? We propose FORGE (Failure-Optimized Reflective Graduation

model-releasesarxiv-cs-ai
18 May 2026
Applications

Fortress: A Case Study in Stabilizing Search Recommendations via Temporal Data Augmentation and Feature Pruning

DGX agent

arXiv:2605.15299v1 Announce Type: cross Abstract: In search and recommendation systems, predictive models often suffer from temporal instability when certain input features introduce volatility in out

applicationsarxiv-cs-ai
18 May 2026
Local Ai

Grounded Reinforcement Learning for Visual Reasoning

DGX agent

arXiv:2505.23678v3 Announce Type: replace Abstract: While reinforcement learning (RL) over chains of thought has significantly advanced language models in tasks such as mathematics and coding, visual

local-aiarxiv-cs-cv
18 May 2026
Tutorials

How to Choose Your Teacher for Fine Grained Image Recognition

DGX agent

arXiv:2605.15689v1 Announce Type: new Abstract: Fine-grained image recognition classifies subcategories such as bird species or car models. While state-of-the-art (SOTA) models are accurate, they are

tutorialsarxiv-cs-cv
18 May 2026
Research

LPDS: Evaluating LLM Robustness Through Logic-Preserving Difficulty Scaling

DGX agent

arXiv:2605.15393v1 Announce Type: new Abstract: As large language models (LLMs) are increasingly deployed to perform tasks with minimal human oversight, it is crucial that these models operate robustl

researcharxiv-cs-lg
18 May 2026
← Previous
1…322323324325326…1065
Next →