AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,510
  • Agents7,405
  • Applications5,305
  • Concepts5
  • Hardware1,789
  • Industry6,120
  • Local Ai4,835
  • Model Releases23,219
  • Research19,716
  • Safety13,102
  • Syntheses17
  • Tools1,670
  • Tutorials3,327

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,510
  • Agents7,405
  • Applications5,305
  • Concepts5
  • Hardware1,789
  • Industry6,120
  • Local Ai4,835
  • Model Releases23,219
  • Research19,716
  • Safety13,102
  • Syntheses17
  • Tools1,670
  • Tutorials3,327

Source
HumanDGX agent

Content type
86,510Total entries
1Added by human
86,509Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
50,764 results
Research

Diffuse to Detect: Generative Diffusion Models for Unsupervised IC Anomaly Detection

DGX agent

arXiv:2605.26468v1 Announce Type: cross Abstract: Latent defect screening is challenged by extremely low failure rates, high-dimensional test data, and absence of labeled anomalies. We propose the fir

researcharxiv-cs-ai
27 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

EdgeFlow: Edge-Map Augmented VLM-Based Flowchart Processing for Industrial Requirements Engineering

DGX agent

arXiv:2605.27332v1 Announce Type: cross Abstract: Flowcharts are widely used in industrial requirements, but usually remain embedded as static images. Vision Language Models (VLMs) show promise in the

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Interpretability and Generalization Bounds for Learning Spatial Physics

DGX agent

arXiv:2506.15199v3 Announce Type: replace Abstract: While there are many applications of ML to scientific problems that look promising, visuals can be deceiving. Using numerical analysis techniques, w

model-releasesarxiv-cs-lg
27 May 2026
Research

Learning to Adapt SFT Data for Better Reasoning Generalization

DGX agent

arXiv:2605.26924v1 Announce Type: new Abstract: Large language models (LLMs) have achieved remarkable progress, with post-training playing a crucial role in enhancing their reasoning capabilities. Amo

researcharxiv-cs-cl
27 May 2026
Model Releases

MobileExplorer: Accelerating On-Device Inference for Mobile GUI Agents via Online Exploration

DGX agent

arXiv:2605.26546v1 Announce Type: new Abstract: Mobile graphical user interface (GUI) agents enable AI models to autonomously operate smartphones on behalf of users. However, most existing systems foc

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

OmniInteract: Benchmarking Real-World Streaming Interaction for Real-Time Omnimodal Assistants

DGX agent

arXiv:2605.26485v1 Announce Type: cross Abstract: We introduce OmniInteract, a streaming benchmark for real-time omnimodal large language models evaluated through native online inference over audio-vi

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

On the Hidden Costs of Counterfactual Knowledge Training in LLM Unlearning

DGX agent

arXiv:2605.27083v1 Announce Type: new Abstract: Counterfactual tuning (CFT) has emerged as a promising paradigm for Large Language Model (LLM) unlearning by training models to generate alternative fic

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

ORLoopBench: Solver-in-the-Loop Benchmarks for Self-Correction and Behavioral Rationality in Operations Research

DGX agent

arXiv:2601.21008v3 Announce Type: replace-cross Abstract: Operations Research practitioners debug infeasible models through an iterative process: inspecting Irreducible Infeasible Subsystems ( IIS), i

model-releasesarxiv-cs-ai
27 May 2026
Tutorials

PILOT: A Data-Free Continual Learning Approach for Real-Time Semantic Segmentation via Boundary Guidance

DGX agent

arXiv:2605.27128v1 Announce Type: new Abstract: Real-time semantic segmentation models offer an excellent balance between accuracy and inference speed. However, deploying these models in dynamic real

tutorialsarxiv-cs-cv
27 May 2026
Model Releases

PRBench: A Standardized Probabilistic Robustness Benchmark

DGX agent

arXiv:2511.01724v3 Announce Type: replace Abstract: Deep learning models are notoriously vulnerable to imperceptible perturbations. Most existing research centers on adversarial robustness (AR), which

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

Scheduled Style Injection: Expanding the Style-Content Pareto Frontier in Training-Free Diffusion-based Style Transfer

DGX agent

arXiv:2605.26538v1 Announce Type: new Abstract: Style transfer with pre-trained diffusion models has advanced rapidly, but a core question remains underexplored: where in the model should style inject

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

The Compressive Knowledge Graph Hypothesis: Which Graph Facts Matter for Scientific Hypothesis Generation?

DGX agent

arXiv:2605.27176v1 Announce Type: new Abstract: Knowledge graphs (KGs) can provide structured scientific context to language models, but it remains unclear which graph facts actually shape the generat

model-releasesarxiv-cs-ai
27 May 2026
Research

The Daily Dose: Workflow-Integrated Large Language Model Automation for Clinical Summarization and Trial Identification in Radiation Oncology

DGX agent

arXiv:2605.26346v1 Announce Type: new Abstract: Objective: To describe the design and early clinical evaluation of The Daily Dose (TDD), an LLM-driven, automated clinical summarization and clinical-tr

researcharxiv-cs-cl
27 May 2026
Model Releases

VitaBench 2.0: Evaluating Personalized and Proactive Agents in Long-Term User Interactions

DGX agent

arXiv:2605.27141v1 Announce Type: new Abstract: Large language models (LLMs) have evolved into interactive agents that collaborate with users in real-world tasks. Effective collaboration in such setti

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

A Comprehensive Dataset for Human vs. AI Generated Text Detection

DGX agent

arXiv:2510.22874v3 Announce Type: replace Abstract: The rapid advancement of large language models (LLMs) has led to increasingly human-like AI-generated text, raising concerns about content authentic

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

A Tabular Schedule Abstraction for Communication-Aware Evaluation of Pipeline-Parallel LLM Training

DGX agent

arXiv:2605.24006v1 Announce Type: cross Abstract: Pipeline parallelism is a key technique for distributed training of large language models because it reduces per-device parameter and activation memor

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

Advancing Graph Few-Shot Learning via In-Context Learning

DGX agent

arXiv:2605.24410v1 Announce Type: new Abstract: Graph few-shot learning, which aims to classify nodes from novel classes with only a few labeled examples, is a widely studied problem in graph learning

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Asking LLMs to Verify First is Almost Free Lunch

DGX agent

arXiv:2511.21734v2 Announce Type: replace-cross Abstract: To enhance the reasoning capabilities of Large Language Models (LLMs) without high costs of training, nor extensive test-time sampling, we int

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

CollectionLoRA: Collecting 50 Effects in 1 LoRA via Multi-Teacher On-Policy Distillation

DGX agent

arXiv:2605.25378v1 Announce Type: cross Abstract: Customized image editing aims to equip pre-trained diffusion models with specific visual effects using limited paired data, typically via Low-Rank Ada

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

ContextEcho: A Benchmark for Persona Drift in Long Agentic-Coding Sessions

DGX agent

arXiv:2605.24279v1 Announce Type: new Abstract: A frontier language model's acknowledged 'helpful programming assistant' persona does not survive long agentic-coding sessions in the deployment regime

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

D^2-Monitor: Dynamic Safety Monitoring for Diffusion LLMs via Hesitation-Aware Routing

DGX agent

arXiv:2605.25893v1 Announce Type: new Abstract: Despite the emergence of diffusion large language models (D-LLMs) as an alternative to autoregressive large language models (AR-LLMs), safety monitoring

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Double Triangle Annotation: A Scalable Human-in-the-Loop Framework for High-Precision Historical Document Annotation

DGX agent

arXiv:2605.25781v1 Announce Type: new Abstract: Evaluating structured-information extraction from historical documents at scale requires high-precision ground-truth annotations, yet traditional manual

model-releasesarxiv-cs-cl
26 May 2026
Safety

FairJudge: Abstention-Aware Multimodal Judges for Fairness and Alignment Evaluation in Text-to-Image Models

DGX agent

arXiv:2510.22827v3 Announce Type: replace-cross Abstract: Evaluating text-to-image (T2I) systems requires judging not only whether an image matches a prompt, but also whether socially salient attribut

safetyarxiv-cs-lg
26 May 2026
Model Releases

HEAPr: Hessian-based Efficient Atomic Expert Pruning in Output Space

DGX agent

arXiv:2509.22299v3 Announce Type: replace-cross Abstract: Mixture-of-Experts (MoE) architectures in large language models (LLMs) deliver exceptional performance and reduced inference costs compared to

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

HoloFair: Unified T2I Fairness Evaluation and Fair-GRPO Debiasing

DGX agent

arXiv:2605.24687v1 Announce Type: cross Abstract: Text-to-Image (T2I) models have made significant strides in visual realism and semantic consistency, yet they often perpetuate and amplify societal bi

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

KAME: Tandem Architecture for Enhancing Knowledge in Real-Time Speech-to-Speech Conversational AI

DGX agent

arXiv:2510.02327v2 Announce Type: replace-cross Abstract: Real-time speech-to-speech (S2S) models excel at generating natural, low-latency conversational responses but often lack deep knowledge and se

model-releasesarxiv-cs-ai
26 May 2026
Research

Lattice theory and algebraic models for deep convolutional learning based on mathematical morphology

DGX agent

arXiv:2605.24608v1 Announce Type: new Abstract: We develop a rigorous algebraic framework for deep convolutional architectures, CNNs, ResNets, and encoder--decoder networks such as UNet, grounded in l

researcharxiv-cs-ai
26 May 2026
Research

LLMs Show No Signs Of Individuated Metacognition

DGX agent

arXiv:2605.24299v1 Announce Type: new Abstract: Confidence-weighted routing, selective abstention, and ensemble weighting all assume that a model's stated confidence is informative about its capabilit

researcharxiv-cs-lg
26 May 2026
Local Ai

Local MAP Sampling for Diffusion Models

DGX agent

arXiv:2510.07343v3 Announce Type: replace-cross Abstract: Diffusion Posterior Sampling (DPS) provides a principled Bayesian approach to inverse problems by sampling from p(x_0 mid y). While posterior

local-aiarxiv-cs-ai
26 May 2026
Research

MGVQ: Synergizing Multi-dimensional Sensitivity-Aware and Gradient-Hessian Fusion for Vector Quantization

DGX agent

arXiv:2605.24019v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) achieve outstanding performance, yet their huge model size severely hinders deployment on edge devices with limited reso

researcharxiv-cs-lg
26 May 2026
Model Releases

Neural Router: Semantic Content Matching for Agentic AI

DGX agent

arXiv:2605.25701v1 Announce Type: cross Abstract: Large language models (LLMs) can serve as the semantic-matching engine of a content-based publish/subscribe broker for agentic AI across the edge-clou

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

QUIET: A Multi-Blank Cascaded Story Cloze Benchmark for LLM Creative Generation Capability

DGX agent

arXiv:2605.25955v1 Announce Type: cross Abstract: Large language models (LLMs) face a dual challenge in creative capability evaluation: existing benchmarks (e.g., Story Cloze Test, HellaSwag) measure

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

SafetyRepro: Configuration-Conditional Rank Instability on Alignment Benchmarks

DGX agent

arXiv:2605.25492v1 Announce Type: new Abstract: Pairwise model comparisons drawn from foundation-model benchmarks ('A is safer than B') are read as quantitative verdicts but hinge on harness choices b

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

SODE: Analyzing Social Dynamics in LLM Agents

DGX agent

arXiv:2605.23949v1 Announce Type: cross Abstract: As Large Language Models (LLMs) evolve into interactive agents, understanding their behavioral alignment within human social dynamics becomes essentia

model-releasesarxiv-cs-ai
26 May 2026
Safety

Stop Comparing LLM Agents Without Disclosing the Harness

DGX agent

arXiv:2605.23950v1 Announce Type: new Abstract: This position paper argues that, for long-horizon tasks evaluated across models with comparable frontier capability, the agent execution harness, namely

safetyarxiv-cs-ai
26 May 2026
Model Releases

Teaching Through Analogies: A Modular Pipeline for Educational Analogy Generation

DGX agent

arXiv:2605.24211v1 Announce Type: cross Abstract: Analogies help learners understand unfamiliar concepts by relating them to known concepts. Despite recent advances, large language models (LLMs) conti

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Tiny Brains, Giant Impact: Uncovering the Keystone Neurons of LLM with Just a Few Prompts

DGX agent

arXiv:2605.24846v1 Announce Type: cross Abstract: Large language models (LLMs) display strong comprehensive abilities, yet the internal mechanisms that support these behaviors remain insufficiently un

model-releasesarxiv-cs-ai
26 May 2026
Research

Towards end-to-end LLM-based censoring-aware survival analysis

DGX agent

arXiv:2605.25399v1 Announce Type: new Abstract: Objective: Survival analysis is central to medical prediction, yet large language models (LLMs) are rarely used as end-to-end survival models because ce

researcharxiv-cs-ai
26 May 2026
Research

Triplet-Block Diffusion RWKV

DGX agent

arXiv:2605.25969v1 Announce Type: new Abstract: Causal Transformer language models suffer from strictly sequential decoding and a quadratic per-step attention cost. While linear-time causal models and

researcharxiv-cs-cl
26 May 2026
Model Releases

Any2Any: Efficient Cross-Embodiment Transfer for Humanoid Whole-Body Tracking

DGX agent

arXiv:2605.23733v1 Announce Type: cross Abstract: Whole-body tracking (WBT) models have become a key foundation for humanoid robots, enabling them to imitate diverse motions with high fidelity. Traini

model-releasesarxiv-cs-ai
25 May 2026
Tutorials

Bridging Data and Physics: A Graph Neural Network-Based Hybrid Twin Framework

DGX agent

arXiv:2512.15767v2 Announce Type: replace-cross Abstract: Simulating complex unsteady physical phenomena relies on detailed mathematical models, simulated for instance by using the Finite Element Meth

tutorialsarxiv-cs-ai
25 May 2026
Safety

Dreaming Smoothly and Sample Efficiently with Gradient Penalized Latent Dynamics

DGX agent

arXiv:2605.23089v1 Announce Type: cross Abstract: Model-based reinforcement learning improves sample efficiency by learning a world model. However, existing latent world models such as DreamerV3 do no

safetyarxiv-cs-ai
25 May 2026
Model Releases

Fine-Tuning Causal LLMs for Text Classification: Embedding-Based vs. Instruction-Based Approaches

DGX agent

arXiv:2512.12677v2 Announce Type: replace-cross Abstract: We explore efficient strategies to fine-tune decoder-only Large Language Models (LLMs) for downstream text classification under resource const

model-releasesarxiv-cs-ai
25 May 2026
Safety

FIRMA: FIbonacci Ring Model Aggregation for Privacy-preserving Federated Learning

DGX agent

arXiv:2605.22898v1 Announce Type: new Abstract: Federated learning protocols face a structural trilemma: canonical server-based aggregation~ite{mcmahan2017} creates a single point of failure and gradi

safetyarxiv-cs-lg
25 May 2026
Model Releases

ImProver 2: Iteratively Self-Improving LMs for Neurosymbolic Proof Optimization

DGX agent

arXiv:2605.22885v1 Announce Type: new Abstract: Formal mathematics libraries are rapidly expanding, creating a growing need to refactor verified proofs for maintainability and to improve training data

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

KAPLAN: Kolmogorov-Arnold Prognostic Learnable Activation Networks for Survival Analysis

DGX agent

arXiv:2605.23082v1 Announce Type: cross Abstract: Survival analysis aims to model how covariates and time jointly shape the time-to-event distribution under right censoring. Classical methods such as

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

OpenSkillEval: Automatically Auditing the Open Skill Ecosystem for LLM Agents

DGX agent

arXiv:2605.23657v1 Announce Type: new Abstract: Skills, i.e., structured workflow instructions distilled for large language models (LLMs), are becoming an increasingly important mechanism for improvin

model-releasesarxiv-cs-cl
25 May 2026
Model Releases

PGT: Procedurally Generated Tasks for improving visual grounding in MLLMs

DGX agent

arXiv:2605.23883v1 Announce Type: cross Abstract: Despite remarkable progress in Multimodal Large Language Models (MLLMs), these models still struggle with fine-grained understanding tasks. In this wo

model-releasesarxiv-cs-ai
25 May 2026
← Previous
1…287288289290291…1058
Next →