AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
HumanDGX agent

Content type
90,914Total entries
1Added by human
90,913Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
53,690 results
Safety

ReasonEdit: Towards Interpretable Image Editing Evaluation via Reinforcement Learning

DGX agent

arXiv:2605.07477v1 Announce Type: new Abstract: Recent text-guided image editing (TIE) models have achieved remarkable progress, however, many edited results still suffer from artifacts, unintended mo

safetyarxiv-cs-cv
11 May 2026
Agents
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

RelAgent: LLM Agents as Data Scientists for Relational Learning

DGX agent

arXiv:2605.07840v1 Announce Type: new Abstract: Relational learning is a challenging problem that has motivated a wide range of approaches, including graph-based models (e.g., graph neural networks, g

agentsarxiv-cs-lg
11 May 2026
Model Releases

ReSeek: A Self-Correcting Framework for Search Agents with Instructive Rewards

DGX agent

arXiv:2510.00568v3 Announce Type: replace Abstract: Search agents powered by Large Language Models (LLMs) have demonstrated significant potential in tackling knowledge-intensive tasks. Reinforcement l

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

RRCM: Ranking-Driven Retrieval over Collaborative and Meta Memories for LLM Recommendation

DGX agent

arXiv:2605.07129v1 Announce Type: cross Abstract: Large Language Models (LLMs) have emerged as a promising paradigm for next-generation recommender systems, offering strong semantic understanding and

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

RuleSafe-VL: Evaluating Rule-Conditioned Decision Reasoning in Vision-Language Content Moderation

DGX agent

arXiv:2605.07760v1 Announce Type: new Abstract: Platform content moderation applies explicit policy rules and context-dependent conditions to decide whether user content is allowed, restricted, or rem

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Scaling Continual Learning to 300+ Tasks with Bi-Level Routing Mixture-of-Experts

DGX agent

arXiv:2602.03473v2 Announce Type: replace-cross Abstract: Continual learning, especially class-incremental learning (CIL), on the basis of a pre-trained model (PTM) has garnered substantial research i

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

SCOPE: Structured Decomposition and Conditional Skill Orchestration for Complex Image Generation

DGX agent

arXiv:2605.08043v1 Announce Type: cross Abstract: While text-to-image models have made strong progress in visual fidelity, faithfully realizing complex visual intents remains challenging because many

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Self-Play Enhancement via Advantage-Weighted Refinement in Online Federated LLM Fine-Tuning with Real-Time Feedback

DGX agent

arXiv:2605.07977v1 Announce Type: new Abstract: Recent works have advanced feedback-based learning systems, whereby a foundation model is able to intake incoming feedback (e.g., a user) to self-improv

model-releasesarxiv-cs-lg
11 May 2026
Research

SHRED: Retain-Set-Free Unlearning via Self-Distillation with Logit Demotion

DGX agent

arXiv:2605.07482v1 Announce Type: cross Abstract: Machine unlearning for large language models (LLMs) aims to selectively remove memorized content such as private data, copyrighted text, or hazardous

researcharxiv-cs-ai
11 May 2026
Research

Stochastic Transition-Map Distillation for Fast Probabilistic Inference

DGX agent

arXiv:2605.07661v1 Announce Type: cross Abstract: Diffusion models achieve strong generation quality, diversity, and distribution coverage, but their performance often comes with expensive inference.

researcharxiv-cs-cv
11 May 2026
Model Releases

Structure Over Scale: Learning Visual Reasoning from Pedagogical Video

DGX agent

arXiv:2601.23251v2 Announce Type: replace Abstract: State-of-the-art vision-language models (VLMs) score impressively on video benchmarks yet stumble on basic visual reasoning tasks involving spatial

model-releasesarxiv-cs-cv
11 May 2026
Local Ai

Teaching Prompts to Coordinate: Hierarchical Layer-Grouped Prompt Tuning for Continual Learning

DGX agent

arXiv:2511.12090v3 Announce Type: replace Abstract: Prompt-based continual learning methods fine-tune only a small set of additional learnable parameters while keeping the pre-trained model's paramete

local-aiarxiv-cs-cv
11 May 2026
Model Releases

Text-to-CAD Evaluation with CADTests

DGX agent

arXiv:2605.07807v1 Announce Type: cross Abstract: Text-to-CAD has recently emerged as an important task with the potential to substantially accelerate design workflows. Despite its significance, there

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

The Translation Tax Is Not a Scalar: A Counterfactual Audit of English-Source Cue Inheritance in Chinese Multilingual Benchmarks

DGX agent

arXiv:2605.07093v1 Announce Type: cross Abstract: The Translation Tax is often treated as a scalar: translated benchmarks are assumed to inflate scores by preserving English-source cues. We audit this

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Tools as Continuous Flow for Evolving Agentic Reasoning

DGX agent

arXiv:2605.07339v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities in orchestrating tools for reasoning tasks. However, existing methods rely on a s

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

TraceAV-Bench: Benchmarking Multi-Hop Trajectory Reasoning over Long Audio-Visual Videos

DGX agent

arXiv:2605.07593v1 Announce Type: new Abstract: Real-world audio-visual understanding requires chaining evidence that is sparse, temporally dispersed, and split across the visual and auditory streams,

model-releasesarxiv-cs-cv
11 May 2026
Tutorials

Transfer Learning Across Fast- and Full-Simulation Domains in High-Energy Physics

DGX agent

arXiv:2605.07471v1 Announce Type: new Abstract: Machine-learning models in high-energy physics are often trained on simulated data, where fully simulated samples are computationally expensive while fa

tutorialsarxiv-cs-lg
11 May 2026
Model Releases

A Comparative Study of PyCaret AutoML and CNN-BiLSTM for Binary Hate Speech Detection in Indonesian Twitter

DGX agent

arXiv:2605.04885v1 Announce Type: new Abstract: This paper compares a PyCaret AutoML branch and a CNN-BiLSTM branch for binary hate speech detection on Indonesian Twitter using the HS label from the c

model-releasesarxiv-cs-cl
7 May 2026
Agents

ARIS: Autonomous Research via Adversarial Multi-Agent Collaboration

DGX agent

arXiv:2605.03042v1 Announce Type: cross Abstract: This report describes ARIS (Auto-Research-in-sleep), an open-source research harness for autonomous research, including its architecture, assurance me

agentsarxiv-cs-ai
7 May 2026
Model Releases

Assessing Cognitive Effort in L2 Idiomatic Processing: An Eye-Tracking Dataset

DGX agent

arXiv:2605.04857v1 Announce Type: new Abstract: This paper presents the development and validation of an eye-tracking dataset designed to investigate how second-language (L2) learners process idiomati

model-releasesarxiv-cs-cl
7 May 2026
Model Releases

AsymmetryZero: A Framework for Operationalizing Human Expert Preferences as Semantic Evals

DGX agent

arXiv:2605.04083v1 Announce Type: new Abstract: Much of the focus in RL today is on evaluation design: building meaningful evals that serve simultaneously as benchmarks and as well-defined reward sign

model-releasesarxiv-cs-lg
7 May 2026
Model Releases

BenCSSmark: Making the Social Sciences Count in LLM Research

DGX agent

arXiv:2605.04886v1 Announce Type: new Abstract: This position paper argues that the under-representation of social science tasks in contemporary LLM benchmarks limits advances in both LLM evaluation a

model-releasesarxiv-cs-cl
7 May 2026
Research

Bilinear Mamba-Koopman Neural MPC for Varying Dynamics

DGX agent

arXiv:2605.04793v1 Announce Type: new Abstract: Koopman-based neural MPC models generate time-varying dynamics from historical data, but preserve convexity by enforcing that the system operator is ind

researcharxiv-cs-lg
7 May 2026
Research

Consensus Entropy: Harnessing Multi-VLM Agreement for Self-Verifying and Self-Improving OCR

DGX agent

arXiv:2504.11101v4 Announce Type: replace Abstract: Optical Character Recognition (OCR) is fundamental to Vision-Language Models (VLMs) and high-quality data generation for LLM training. Yet, despite

researcharxiv-cs-cv
7 May 2026
Model Releases

ConsisVLA-4D: Advancing Spatiotemporal Consistency in Efficient 3D-Perception and 4D-Reasoning for Robotic Manipulation

DGX agent

arXiv:2605.05126v1 Announce Type: new Abstract: Current Vision-Language-Action (VLA) models primarily focus on mapping 2D observations to actions, but exhibit notable limitations in spatiotemporal per

model-releasesarxiv-cs-ro
7 May 2026
Model Releases

DALight-3D: A Lightweight 3D U-Net for Brain Tumor Segmentation from Multi-Modal MRI

DGX agent

arXiv:2605.04518v1 Announce Type: new Abstract: Automatic brain tumor segmentation from multi-modal MRI remains challenging because volumetric models often incur substantial computational cost. This p

model-releasesarxiv-cs-cv
7 May 2026
Research

Dataset-Driven Channel Masks in Transformers for Multivariate Time Series

DGX agent

arXiv:2410.23222v3 Announce Type: replace Abstract: Recent advancements in foundation models have been successfully extended to the time series (TS) domain, facilitated by the emergence of large-scale

researcharxiv-cs-lg
7 May 2026
Model Releases

Discovering New Theorems via LLMs with In-Context Proof Learning in Lean

DGX agent

arXiv:2509.14274v2 Announce Type: replace Abstract: Large Language Models (LLMs) have demonstrated significant promise in formal theorem proving. In this study, we investigate the ability of LLMs to d

model-releasesarxiv-cs-lg
7 May 2026
Safety

Elicitation Matters: How Prompts and Query Protocols Shape LLM Surrogates under Sparse Observations

DGX agent

arXiv:2605.04764v1 Announce Type: new Abstract: Large language models are increasingly used as surrogate models for low-data optimization, but their optimizer-facing prediction and its uncertainty rem

safetyarxiv-cs-cl
7 May 2026
Applications

ELVIS: Ensemble-Calibrated Latent Imagination for Long-Horizon Visual MPC

DGX agent

arXiv:2605.04709v1 Announce Type: new Abstract: A central challenge of visual control with model-based reinforcement learning (RL) is reliable long-horizon planning: long rollouts with learned latent

applicationsarxiv-cs-lg
7 May 2026
Model Releases

Empirical Study of Pop and Jazz Mix Ratios for Genre-Adaptive Chord Generation

DGX agent

arXiv:2605.04998v1 Announce Type: cross Abstract: Chord progression generation is practically important but understudied. Most large-scale symbolic music systems target melody, multi-track arrangement

model-releasesarxiv-cs-lg
7 May 2026
Model Releases

Explaining and Preventing Alignment Collapse in Iterative RLHF

DGX agent

arXiv:2605.04266v1 Announce Type: new Abstract: Reinforcement learning from human feedback (RLHF) typically assumes a static or non-strategic reward model (RM). In iterative deployment, however, the p

model-releasesarxiv-cs-lg
7 May 2026
Model Releases

Feature Identification via the Empirical NTK

DGX agent

arXiv:2510.00468v4 Announce Type: replace Abstract: We provide evidence that eigenanalysis of the empirical neural tangent kernel (eNTK) can surface feature directions in trained neural networks. Acro

model-releasesarxiv-cs-lg
7 May 2026
Model Releases

Fragile Knowledge, Robust Instruction-Following: The Width Pruning Dichotomy in Llama-3.2

DGX agent

arXiv:2512.22671v2 Announce Type: replace Abstract: Structured width pruning of GLU-MLP layers, guided by the Maximum Absolute Weight (MAW) criterion, reveals a systematic dichotomy in how reducing th

model-releasesarxiv-cs-cl
7 May 2026
Model Releases

Generalization Bounds of Spiking Neural Networks via Rademacher Complexity

DGX agent

arXiv:2605.02927v1 Announce Type: cross Abstract: Spiking Neural Networks (SNNs) have garnered increasing attention as one of bio-inspired models due to their great potential in neuromorphic computing

model-releasesarxiv-cs-ai
7 May 2026
Model Releases

Generative Quantum-inspired Kolmogorov-Arnold Eigensolver

DGX agent

arXiv:2605.04604v1 Announce Type: cross Abstract: High-performance computing (HPC) is increasingly important for scalable quantum chemistry workflows that couple classical generative models, quantum c

model-releasesarxiv-cs-lg
7 May 2026
Model Releases

Intermediate Representations are Strong AI-Generated Image Detectors

DGX agent

arXiv:2605.04358v1 Announce Type: new Abstract: The rapid advancement in generative AI models has enabled the creation of photorealistic images. At the same time, there are growing concerns about the

model-releasesarxiv-cs-cv
7 May 2026
Model Releases

MP-ISMoE: Mixed-Precision Interactive Side Mixture-of-Experts for Efficient Transfer Learning

DGX agent

arXiv:2605.04058v1 Announce Type: new Abstract: Parameter-efficient transfer learning (PETL) has emerged as a pivotal paradigm for adapting pre-trained foundation models to downstream tasks, significa

model-releasesarxiv-cs-lg
7 May 2026
Model Releases

OpenVTON-Bench: A Large-Scale High-Resolution Benchmark for Controllable Virtual Try-On Evaluation

DGX agent

arXiv:2601.22725v3 Announce Type: replace Abstract: Recent advances in diffusion models have significantly elevated the visual fidelity of Virtual Try-On (VTON) systems, yet reliable evaluation remain

model-releasesarxiv-cs-cv
7 May 2026
Safety

OracleProto: A Reproducible Framework for Benchmarking LLM Native Forecasting via Knowledge Cutoff and Temporal Masking

DGX agent

arXiv:2605.03762v1 Announce Type: new Abstract: Large language models are moving from static text generators toward real-world decision-support systems, where forecasting is a composite capability tha

safetyarxiv-cs-ai
7 May 2026
Model Releases

Transformation Categorization Based on Group Decomposition Theory Using Parameter Division

DGX agent

arXiv:2605.04056v1 Announce Type: new Abstract: Representation learning seeks meaningful sensory representations without supervision and can model aspects of human development. Although many neural ne

model-releasesarxiv-cs-lg
7 May 2026
Model Releases

UFAL-CUNI at SemEval-2026 Task 11: An Efficient Modular Neuro-symbolic Method for Syllogistic Reasoning

DGX agent

arXiv:2605.04941v1 Announce Type: new Abstract: This paper describes our system submitted to SemEval-2026 Task 11: Disentangling Content and Formal Reasoning in Large Language Models. We present an ef

model-releasesarxiv-cs-cl
7 May 2026
Model Releases

AI and Open-data Driven Scalable Solar Power Profiling

DGX agent

arXiv:2605.02738v1 Announce Type: new Abstract: Solar photovoltaic (PV) deployment is expanding rapidly, yet detailed, up-to-date information on the spatial distribution and capacity of rooftop PV rem

model-releasesarxiv-cs-ai
6 May 2026
Model Releases

Amortized Variational Inference for Joint Posterior and Predictive Distributions in Bayesian Uncertainty Quantification

DGX agent

arXiv:2605.03710v1 Announce Type: cross Abstract: Bayesian predictive inference propagates parameter uncertainty to quantities of interest through the posterior-predictive distribution. In practice, t

model-releasesarxiv-cs-lg
6 May 2026
Model Releases

Artificial Jagged Intelligence as Uneven Optimization Energy Allocation Capability Concentration, Redistribution, and Optimization Governance

DGX agent

arXiv:2605.01420v1 Announce Type: new Abstract: Artificial Jagged Intelligence (AJI) denotes a recurring pattern in which large learning systems exhibit strong local capabilities while remaining weak

model-releasesarxiv-cs-ai
6 May 2026
Model Releases

Benchmarking Document Parsers on Mathematical Formula Extraction from PDFs

DGX agent

arXiv:2512.09874v2 Announce Type: replace Abstract: Correctly parsing mathematical formulas from PDFs is critical for training large language models and building scientific knowledge bases from academ

model-releasesarxiv-cs-cv
6 May 2026
Model Releases

Calibration of the underlying surface parameters for urban flood using latent variables and adjoint equation

DGX agent

arXiv:2605.02959v1 Announce Type: new Abstract: Calibrating the urban underlying surface parameters is crucial for urban flood simulation. We formulate the parameter calibration problem into an optimi

model-releasesarxiv-cs-lg
6 May 2026
Model Releases

ContextCov: Deriving and Enforcing Executable Constraints from Agent Instruction Files

DGX agent

arXiv:2603.00822v2 Announce Type: replace-cross Abstract: As Large Language Model (LLM) agents increasingly execute complex, autonomous software engineering tasks, developers rely on natural language

model-releasesarxiv-cs-ai
6 May 2026
← Previous
1…451452453454455…1119
Next →