AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,042
  • Agents7,457
  • Applications5,325
  • Concepts5
  • Hardware1,802
  • Industry6,143
  • Local Ai4,863
  • Model Releases23,393
  • Research19,837
  • Safety13,180
  • Syntheses17
  • Tools1,671
  • Tutorials3,349

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,042
  • Agents7,457
  • Applications5,325
  • Concepts5
  • Hardware1,802
  • Industry6,143
  • Local Ai4,863
  • Model Releases23,393
  • Research19,837
  • Safety13,180
  • Syntheses17
  • Tools1,671
  • Tutorials3,349

Source
HumanDGX agent

Content type
87,042Total entries
1Added by human
87,041Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,106 results
Model Releases

Attention Capture Is Not Detection: A Two-Stage Account of How Humans Miss Localized AI Image Edits

DGX agent

arXiv:2608.13865v1 Announce Type: new Abstract: As AI-generated image edits proliferate, the platforms meant to curb the resulting disinformation treat detectability as a single, undifferentiated prop

model-releasesarxiv-cs-cv
17 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Buy the Rumor, Sell the News: When Is News Priced In?

DGX agent

arXiv:2608.14014v1 Announce Type: new Abstract: Two old market sayings hold that news is already priced in by the time it is published, and that the rumor is bought while the news is sold. Both place

model-releasesarxiv-cs-ai
17 Aug 2026
Safety

CForce: Boosting Parallel Decoding for dLLMs via Consistency Forcing

DGX agent

arXiv:2608.13925v1 Announce Type: cross Abstract: Diffusion large language models (dLLMs) accelerate language generation by predicting multiple masks in a single forward pass. However, existing dLLMs

safetyarxiv-cs-ai
17 Aug 2026
Model Releases

CVT-Bench: Probing Spatial-State Integrity through Counterfactual Viewpoint Transformations

DGX agent

arXiv:2603.21114v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) perform strongly on isolated spatial tasks, but whether their predictions remain persistent and mutually co

model-releasesarxiv-cs-cv
17 Aug 2026
Model Releases

Evaluating Agentic Learning Harness Capabilities Without Labels via the Scaling Hypothesis

DGX agent

arXiv:2608.13608v1 Announce Type: new Abstract: Agentic 'Continual Learning Harnesses', systems that pair an LLM with retrieval or memory to improve from feedback without retraining, have shown growin

model-releasesarxiv-cs-ai
17 Aug 2026
Model Releases

LightTeaNet: A Weakly Supervised Lightweight CNN for Multi-Label Tea Leaf Disease Detection and Localization

DGX agent

arXiv:2608.14178v1 Announce Type: new Abstract: Tea is known as an important crop in many parts of South and Southeast Asia, yet the production of tea is still hampered by the multiple diseases that d

model-releasesarxiv-cs-cv
17 Aug 2026
Model Releases

MACS: A Hybrid Multi-Agent Framework for Reliable Conversational E-Commerce Recommendation

DGX agent

arXiv:2608.14068v1 Announce Type: cross Abstract: Conversational recommendation for e-commerce is increasingly mediated by large language models (LLMs), yet many real-world deployments operate under a

model-releasesarxiv-cs-ai
17 Aug 2026
Local Ai

Not All Tokens Are Equal: Inflation-Aware Routing for Agentic LLM Systems

DGX agent

arXiv:2608.13571v1 Announce Type: cross Abstract: When a language model fails to answer a query on the first attempt, an agentic system retries, consuming additional tokens each time. This retry overh

local-aiarxiv-cs-ai
17 Aug 2026
Research

On the Brittleness of Maximum Likelihood Estimation for Gaussian Process Hyperparameter Optimization

DGX agent

arXiv:2608.13793v1 Announce Type: cross Abstract: Machine learning (ML) has become an indispensable part of modern engineering design workflows. A crucial step in training an ML model is the selection

researcharxiv-cs-lg
17 Aug 2026
Safety

Responsiveness Verification: Will Predictions Change? How Much? How Often?

DGX agent

arXiv:2507.02169v2 Announce Type: replace Abstract: Machine learning models are often used in applications where their inputs change due to routine interactions, strategic manipulation, or noise. In s

safetyarxiv-cs-lg
17 Aug 2026
Model Releases

Wrong but Useful: Trajectory Value Beyond Answer Correctness in Multi-Agent Messages

DGX agent

arXiv:2608.14375v1 Announce Type: new Abstract: Multi-agent reasoning systems often use agreement, confidence, or automated scores to decide which messages should shape a final answer. Such filtering

model-releasesarxiv-cs-ai
17 Aug 2026
Model Releases

A Generative Approach for Improving Multi-Label Defect Classification in Photovoltaic Modules

DGX agent

arXiv:2608.12725v1 Announce Type: new Abstract: This paper addresses the challenge of multi-label defect classification in electroluminescence (EL) images of photovoltaic (PV) cells. Training models o

model-releasesarxiv-cs-cv
14 Aug 2026
Research

Are You Sure You're Sure? On the Impact of Instruction Tuning on Confidence and Lexical Diversity

DGX agent

arXiv:2608.13430v1 Announce Type: cross Abstract: Instruction-tuned language models achieve strong performance across a range of generation tasks, but have also recently been shown to exhibit verbaliz

researcharxiv-cs-ai
14 Aug 2026
Model Releases

AutoDesign: Meta-Harness Optimization for Long-Horizon Agentic Design

DGX agent

arXiv:2608.13560v1 Announce Type: cross Abstract: Transforming multimodal sources into condensed and structured media outputs can be fundamentally conceptualized as a long-horizon agentic process cent

model-releasesarxiv-cs-ai
14 Aug 2026
Model Releases

CoMedBench: A Multi-Source Benchmark of Synthetic Medical Data Fidelity and Downstream Utility

DGX agent

arXiv:2608.12805v1 Announce Type: new Abstract: Access to clinical data is essential for developing reliable healthcare machine learning systems, but direct use of electronic health records is constra

model-releasesarxiv-cs-lg
14 Aug 2026
Research

Dead text or binding clause? Measuring and restoring constraint influence in black-box LLM dialogues

DGX agent

arXiv:2608.12599v1 Announce Type: new Abstract: Multi-turn dialogues let users revoke constraints as easily as impose them, but revocation does not reliably take effect: models keep enacting withdrawn

researcharxiv-cs-ai
14 Aug 2026
Research

EEG-PRIME: Prototype-Aligned Representation Learning with Multi-Level Conditioning for EEG Decoding

DGX agent

arXiv:2608.13072v1 Announce Type: new Abstract: Electroencephalography (EEG) decoding models often generalize poorly across datasets and subjects due to domain shifts in acquisition protocols and indi

researcharxiv-cs-ai
14 Aug 2026
Local Ai

From Local Mismatch to Global Impact: Optimizing Cache Reuse Policy for Efficient Diffusion

DGX agent

arXiv:2608.13043v1 Announce Type: new Abstract: Diffusion models have achieved dominant performance in visual generation but suffer from substantial inference overhead. While cache-based acceleration

local-aiarxiv-cs-ai
14 Aug 2026
Research

Geometric and Behavioral Stratification in Transformer Residual Streams

DGX agent

arXiv:2608.12447v1 Announce Type: cross Abstract: Trained transformer models develop privileged bases: coordinate axes whose statistics differ from the rest of the residual stream. But what kind of di

researcharxiv-cs-cl
14 Aug 2026
Applications

Incremental Evaluation and Training in Relational Deep Learning

DGX agent

arXiv:2608.13023v1 Announce Type: new Abstract: Relational Deep Learning (RDL) models multi-tabular databases as temporal heterogeneous graphs to enable end-to-end representation learning. However, pr

applicationsarxiv-cs-lg
14 Aug 2026
Safety

Jagged Judges: Epistemic Stability Under Silence, Pressure, and Persistence

DGX agent

arXiv:2608.12645v1 Announce Type: new Abstract: LLM judges have become central infrastructure for model evaluations, online grading, and reward modeling. Judges are typically validated by accuracy on

safetyarxiv-cs-ai
14 Aug 2026
Research

Keep, Customize, or Exit: Default Design and Token Pricing in LLM Reasoning Services

DGX agent

arXiv:2608.13315v1 Announce Type: cross Abstract: We study a large language model (LLM) service in which a provider chooses a per-token price and a default reasoning-token allocation, while a user may

researcharxiv-cs-ai
14 Aug 2026
Research

Masked diffusion LLMs can use EoS tokens for hidden reasoning

DGX agent

arXiv:2603.05197v2 Announce Type: replace Abstract: Diffusion LLMs have been proposed as an alternative to autoregressive LLMs. Curiously, they are especially capable if the generation length, i.e., t

researcharxiv-cs-cl
14 Aug 2026
Tutorials

New Terms, New Toxicity: Consensus-based Chinese Neologism Toxicity Detection via Search-Augmented LLMs

DGX agent

arXiv:2608.12361v1 Announce Type: new Abstract: Neologisms, emerging terms in meaning or form, can serve as new vehicles for toxic expression, like 'country girl' as a stigmatizing label targeting fem

tutorialsarxiv-cs-cl
14 Aug 2026
Model Releases

Predicting consumer-technology ownership without a diffusion history

DGX agent

arXiv:2608.12344v1 Announce Type: new Abstract: We test whether the perceived attributes of a consumer technology predict how widely it is owned. In a 2022 Prolific survey of US adults (n = 678), resp

model-releasesarxiv-cs-cl
14 Aug 2026
Model Releases

Reasoning for Social Audio-Visual Question Answering: Where Do We Stand?

DGX agent

arXiv:2608.13239v1 Announce Type: new Abstract: Training Multimodal Large Language Models for audio-visual social understanding is a crucial step toward embodied social intelligence. Chain-of-thought

model-releasesarxiv-cs-cv
14 Aug 2026
Model Releases

Reliability-Aware Sexism Detection: Combining DPO with Annotator Agreement and Token-Level Confidence Scoring

DGX agent

arXiv:2608.12330v1 Announce Type: new Abstract: The detection of online sexism remains an open problem. Sexism detection is inherently subjective, yet most existing systems reduce multi-annotator labe

model-releasesarxiv-cs-cl
14 Aug 2026
Model Releases

RoboSynChallenge: Mastering Real-World Dexterity via Generalizing Synthesized Manipulation Skills

DGX agent

arXiv:2608.12416v1 Announce Type: new Abstract: Achieving generalizable robotic manipulation remains a central challenge in embodied intelligence. Despite rapid advances in model architectures and lea

model-releasesarxiv-cs-ro
14 Aug 2026
Model Releases

TEMPO: Makespan-Aware Expert-Parallel Load Balancing Across Memory- and Compute-Bound Regimes

DGX agent

arXiv:2608.13057v1 Announce Type: cross Abstract: In expert-parallel (EP) MoE serving, every layer synchronizes at the slowest GPU. Dispatchers balance token counts (EPLB, LPLB, UltraEP) or activated-

model-releasesarxiv-cs-ai
14 Aug 2026
Model Releases

Toward a Gricean Retreat: Probing LLMs for Knowledge Boundaries and Referent Specificity

DGX agent

arXiv:2608.13484v1 Announce Type: cross Abstract: When asked about entities outside their knowledge boundary, LLMs routinely fabricate plausible-sounding details rather than backing off to safer, more

model-releasesarxiv-cs-ai
14 Aug 2026
Model Releases

Unifying Depth and Width Pruning for LLMs via Binary Knapsack Optimization

DGX agent

arXiv:2608.12953v1 Announce Type: new Abstract: Structured pruning is a promising approach for compressing large language models (LLMs), yet existing methods rely heavily on greedy heuristics that pro

model-releasesarxiv-cs-cl
14 Aug 2026
Model Releases

Where You Measure Decides What You Measure: Position Selection in Ablation-Based SAE Evaluation

DGX agent

arXiv:2608.13337v1 Announce Type: new Abstract: Sparse autoencoders are meant to name the things a language model computes, and the usual way to check that a latent matters is to switch it off and see

model-releasesarxiv-cs-lg
14 Aug 2026
Model Releases

A Neighborhood Attention Transformer Network for Enhanced 3D Segmentation of the Left Anterior Descending Artery

DGX agent

arXiv:2608.12274v1 Announce Type: cross Abstract: Background: Accurate segmentation of the Left Anterior Descending (LAD) artery in 3D free-breathing, non-contrast CT is critical for cardiac dose spar

model-releasesarxiv-cs-ai
13 Aug 2026
Model Releases

Advancing MLLM-based UAV Image Understanding and Reasoning: A Benchmark and a Training-Free Multi-Agent System

DGX agent

arXiv:2608.11738v1 Announce Type: cross Abstract: Multimodal Large Language Model (MLLM)-based UAV aerial image understanding and reasoning is essential for aerial intelligence yet poses distinct chal

model-releasesarxiv-cs-ai
13 Aug 2026
Model Releases

Agent Safety Should Be a Runtime Contract

DGX agent

arXiv:2608.11274v1 Announce Type: cross Abstract: The dominant paradigm treats AI safety as a property to be instilled during model training via RLHF, DPO, or Constitutional AI. We argue this is struc

model-releasesarxiv-cs-ai
13 Aug 2026
Research

APEX: Adaptive Expert Prefetching for Memory-Efficient Edge MoE Inference

DGX agent

arXiv:2608.11688v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) models are attractive for edge deployment because they provide high model capacity while activating only a small subset of pa

researcharxiv-cs-ai
13 Aug 2026
Model Releases

Benchmarking Cyberattack Detection in Electric Vehicle Charging Infrastructure with Benign User Updates

DGX agent

arXiv:2608.11286v1 Announce Type: cross Abstract: Cyberattack detection in electric vehicle charging infrastructure is complicated by legitimate post-activation revisions to requested energy and depar

model-releasesarxiv-cs-lg
13 Aug 2026
Local Ai

Diffusion-Based Data-Driven Assortment Optimization

DGX agent

arXiv:2608.11419v1 Announce Type: new Abstract: Assortment optimization is a fundamental problem in revenue management, typically addressed using parametric choice models such as the multinomial logit

local-aiarxiv-cs-lg
13 Aug 2026
Model Releases

DREAMS: Density Functional Theory Based Research Engine for Agentic Materials Simulation

DGX agent

arXiv:2507.14267v2 Announce Type: replace Abstract: Large language model (LLM) agents can execute long-horizon scientific workflows, but their numerical outputs are difficult to trust: agents lose con

model-releasesarxiv-cs-ai
13 Aug 2026
Research

GeoFlow: Efficient Driving Video Generation via Geometry-Aligned Priors

DGX agent

arXiv:2608.12203v1 Announce Type: new Abstract: Generative models like Diffusion Models and Flow Matching have demonstrated remarkable capabilities in synthesizing high-fidelity driving videos, but ar

researcharxiv-cs-cv
13 Aug 2026
Research

LEMUR: Latent Entropy-aware Multimodal Unlearning via Visual-anchored Reasoning Redirection

DGX agent

arXiv:2608.11691v1 Announce Type: cross Abstract: Reinforcement-learning (RL) post-training equips multimodal large reasoning models (MLRMs) with exploratory chains of thought (CoT), substantially imp

researcharxiv-cs-cl
13 Aug 2026
Model Releases

Lifecycle-Optimal Tokenization: Vocabulary Size as a Deployment-Regime-Dependent Infrastructure Parameter

DGX agent

arXiv:2608.11361v1 Announce Type: cross Abstract: Tokenizer vocabulary size is a foundational design choice in large language model (LLM) infrastructure, yet it is typically fixed at training time bas

model-releasesarxiv-cs-cl
13 Aug 2026
Model Releases

Located but Not Releasable: Silent Gate Inversion and Bounded Linear Release

DGX agent

arXiv:2608.11822v1 Announce Type: new Abstract: A growing body of work reports that language models represent task-relevant latent structure that they fail to use. Whether such structure, once located

model-releasesarxiv-cs-cl
13 Aug 2026
Applications

Look What the Probes Dragged In! Real-World Chest X-ray Shortcuts in MedCLIP

DGX agent

arXiv:2608.12086v1 Announce Type: new Abstract: Vision-language models, such as contrastive language-image pre-training (CLIP)-based approaches, have reached state-of-the-art (SOTA) results in medical

applicationsarxiv-cs-cv
13 Aug 2026
Model Releases

LoRAQuant: Mixed-Precision Quantization of LoRA to Ultra-Low Bits

DGX agent

arXiv:2510.26690v3 Announce Type: replace Abstract: Low-Rank Adaptation (LoRA) has become a popular technique for parameter-efficient fine-tuning of large language models (LLMs). In many real-world sc

model-releasesarxiv-cs-lg
13 Aug 2026
Safety

Quantifying the Relationship Between Clinical Safety and Environmental Impact in Therapeutic LLMs

DGX agent

arXiv:2608.11830v1 Announce Type: cross Abstract: The deployment of large language models (LLMs) in mental health contexts raises questions about the relationship between clinical safety and environme

safetyarxiv-cs-cl
13 Aug 2026
Model Releases

SoftWater: Class-Aware Rate Allocation for Softmax Quantization

DGX agent

arXiv:2608.12026v1 Announce Type: new Abstract: Post-training quantization pipelines routinely leave the softmax output layer in high precision. Yet in small LLMs with modern vocabularies, the head ho

model-releasesarxiv-cs-lg
13 Aug 2026
Model Releases

Benchmarking LLM-Guided Control-Plane Policies for Backend Fault Isolation in HAProxy

DGX agent

arXiv:2608.10532v1 Announce Type: cross Abstract: Static load balancers cannot mitigate a backend that is degraded rather than down: round-robin and least-connections keep routing traffic to a server

model-releasesarxiv-cs-lg
12 Aug 2026
← Previous
1…336337338339340…1065
Next →