AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,457
  • Agents7,399
  • Applications5,302
  • Concepts5
  • Hardware1,786
  • Industry6,117
  • Local Ai4,835
  • Model Releases23,193
  • Research19,715
  • Safety13,094
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,457
  • Agents7,399
  • Applications5,302
  • Concepts5
  • Hardware1,786
  • Industry6,117
  • Local Ai4,835
  • Model Releases23,193
  • Research19,715
  • Safety13,094
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
86,457Total entries
1Added by human
86,456Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
50,764 results
Model Releases

HyperODE: Zero-Shot Surrogate for Simulation and Inference of Dynamical Systems

DGX agent

arXiv:2608.00852v1 Announce Type: new Abstract: Understanding and controlling complex dynamical systems often requires executing thousands of numerical simulations across vast parametric landscapes, w

model-releasesarxiv-cs-lg
4 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

It's the Decoding Format, Not the Perturbation: Auditing Consistency-Based Selection for Vision-Language Test-Time Scaling

DGX agent

arXiv:2608.01207v1 Announce Type: new Abstract: Test-time scaling lifts large language model reasoning by sampling many candidate solutions and selecting among them, yet the same recipe transfers poor

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

Light Forcing: Accelerating Autoregressive Video Diffusion via Sparse Attention

DGX agent

arXiv:2602.04789v4 Announce Type: replace Abstract: Advanced autoregressive (AR) video generation models have improved visual fidelity and interactivity, but the quadratic complexity of attention rema

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

MoRAL: Sensor-Grounded BEV Reasoning for Compact VLMs toward Edge-Oriented Autonomous Driving

DGX agent

arXiv:2608.02449v1 Announce Type: new Abstract: Deploying vision-language models (VLMs) for safety-critical spatial reasoning on resource-constrained autonomous driving platforms requires both compact

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

PredAct-Bench: Benchmarking Tool-Augmented Dialogue under Controlled Tool Noise

DGX agent

arXiv:2608.02372v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed in task-oriented dialogue systems that support multi-step decision-making in high-stakes domains

model-releasesarxiv-cs-cl
4 Aug 2026
Safety

Proteus: A Truncation-Robust Entropy Model for Progressive LiDAR Compression

DGX agent

arXiv:2608.00687v1 Announce Type: new Abstract: LiDAR point clouds provide explicit, deterministic physical boundaries critical for collaborative safety-critical perception. However, wireless channels

safetyarxiv-cs-cv
4 Aug 2026
Model Releases

Refine Drugs, Don't Complete Them: Uniform-Source Discrete Flows for Fragment-Based Drug Discovery

DGX agent

arXiv:2509.26405v2 Announce Type: replace Abstract: We introduce InVirtuoGen, a discrete flow generative model for fragmented SMILES for de novo and fragment-constrained generation, and target-propert

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

Similarity-Aware Machine Unlearning

DGX agent

arXiv:2608.00246v1 Announce Type: new Abstract: Machine unlearning removes the influence of user-specified training examples from a trained model, avoiding the need to retrain it from scratch. Localiz

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

Toward Robust LLM-Based Judges: Taxonomic Bias Evaluation and Debiasing Optimization

DGX agent

arXiv:2603.08091v2 Announce Type: replace Abstract: Large language model (LLM)-based judges are widely adopted for automated evaluation and reward modeling, yet their judgments are often affected by j

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

TreeProbe : A Tibetan Medicine Benchmark for Cultural Bias in LLMs

DGX agent

arXiv:2608.00640v1 Announce Type: new Abstract: Large language models are increasingly viewed as a potential means of mitigating global health inequities, yet their outputs often reflect dominant high

model-releasesarxiv-cs-cl
4 Aug 2026
Safety

Who Should Be Generated? Justifying Demographic Targets in Open-Ended Generation

DGX agent

arXiv:2608.02551v1 Announce Type: cross Abstract: Fairness evaluation concerns not only what a model produces, but also what its outputs ought to be compared against. When a model generates 'a CEO in

safetyarxiv-cs-cl
4 Aug 2026
Research

A Generalized-Bayes Perspective on Counterfactual Explanations: Posterior-Based Decision-Making and Evaluation

DGX agent

arXiv:2607.29077v1 Announce Type: new Abstract: Counterfactual explanations (CEs) enhance the interpretability of machine learning models by identifying the smallest change to an input required to obt

researcharxiv-cs-ai
3 Aug 2026
Model Releases

Agentic Harness for Real-World Compilers

DGX agent

arXiv:2603.20075v2 Announce Type: replace-cross Abstract: Compilers are critical to modern computing, yet fixing compiler bugs is difficult. While recent large language model (LLM) advancements enable

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

Benchmarks Are Not Monolithic: Sample-Level Auditing and Orchestration for LLM Evaluation

DGX agent

arXiv:2607.28801v1 Announce Type: cross Abstract: Benchmark datasets are central to evaluating Large Language Models (LLMs), yet they are typically conceived as monolithic tasks, obscuring substantial

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

Benchmarks Are Not Validation: A System-Level View of Financial LLM Applications

DGX agent

arXiv:2607.28840v1 Announce Type: new Abstract: Large language models are increasingly deployed in financial applications that combine retrieval, proprietary data, tool use, orchestration logic, monit

model-releasesarxiv-cs-cl
3 Aug 2026
Research

Can Synthetic Data Overcome the Generalization Limits of AI-Based Flower and Pod Detection Across Cowpea Breeding Genotypes and Environments?

DGX agent

arXiv:2607.28796v1 Announce Type: new Abstract: High-throughput phenotyping requires AI-enabled computer vision models that generalize across genotypes, locations, and growing seasons, yet such models

researcharxiv-cs-cv
3 Aug 2026
Agents

HarnessBank: Semantic Gene-Bank Search with Gated Verification for Agent-Harness Self-Evolution

DGX agent

arXiv:2607.13683v2 Announce Type: replace Abstract: Large Language Models (LLMs) have enabled capable agents across diverse applications. Beyond the foundation model, the performance of an agent is go

agentsarxiv-cs-cl
3 Aug 2026
Model Releases

LARA: Lightweight Adapters in the Residual Stream for Composable Adaptation and Alignment

DGX agent

arXiv:2607.28669v1 Announce Type: new Abstract: We present LARA (Lightweight Additive Residual Adaptation), a method for efficient adaptation that operates in the residual stream of a frozen model rat

model-releasesarxiv-cs-lg
3 Aug 2026
Local Ai

Multimodal Reinforcement Learning with Adaptive Verifier for AI Agents

DGX agent

arXiv:2512.03438v3 Announce Type: replace Abstract: Agentic reasoning models trained with multimodal reinforcement learning (MMRL) have become increasingly capable, yet they are almost universally opt

local-aiarxiv-cs-ai
3 Aug 2026
Applications

Pay for The Second-Best Service: A Game-Theoretic Approach Against Dishonest LLM Providers

DGX agent

arXiv:2511.00847v5 Announce Type: replace-cross Abstract: The widespread adoption of Large Language Models (LLMs) through Application Programming Interfaces (APIs) induces a critical vulnerability: th

applicationsarxiv-cs-ai
3 Aug 2026
Local Ai

Small Is Enough: Per-User Style Rewriting of AI-Edited Text via LoRA Adapters

DGX agent

arXiv:2607.29238v1 Announce Type: cross Abstract: InMyStyle is a privacy first, single user system that adapts small language models to rewrite AI-edited text towards an individual user's writing styl

local-aiarxiv-cs-ai
3 Aug 2026
Model Releases

So-Fake: Benchmarking and Explaining Social Media Image Forgery Detection

DGX agent

arXiv:2505.18660v5 Announce Type: replace Abstract: Recent advances in AI-powered generative models have enabled the creation of increasingly realistic synthetic images, posing significant risks to in

model-releasesarxiv-cs-cv
3 Aug 2026
Model Releases

WebCoderBench: Benchmarking Web Application Generation with Comprehensive and Interpretable Evaluation Metrics

DGX agent

arXiv:2601.02430v3 Announce Type: replace-cross Abstract: Web applications (web apps) have become a key arena for large language models (LLMs) to demonstrate their code generation capabilities and com

model-releasesarxiv-cs-ai
3 Aug 2026
Applications

ARES: Anomaly Recognition Model For Edge Streams

DGX agent

arXiv:2511.22078v2 Announce Type: replace Abstract: Many real-world scenarios involving streaming information can be represented as temporal graphs, where data flows through dynamic changes in edges o

applicationsarxiv-cs-lg
31 Jul 2026
Model Releases

Benchmarking LLM Competence on Logical Inference over Probability Operators

DGX agent

arXiv:2607.27405v1 Announce Type: new Abstract: Both expressions of uncertainty and inferences are ubiquitous in natural language, and valid inferences over natural-language expressions of uncertainty

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

Beyond Borrowed Histories: Person-Aligned User Simulation for Interactive Role-Playing Evaluation

DGX agent

arXiv:2607.27816v1 Announce Type: new Abstract: Role-playing agents (RPAs) have become one of the most important consumer applications of large language models. Users engage in multi-turn conversation

model-releasesarxiv-cs-cl
31 Jul 2026
Research

Calibrate Before Reason: Robust Visual Token Reduction against Semantic Drift in VLMs

DGX agent

arXiv:2607.27700v1 Announce Type: new Abstract: Large Vision-Language Models (VLMs) suffer from prohibitive inference overhead due to long sequences of visual tokens. However, existing visual token re

researcharxiv-cs-cv
31 Jul 2026
Applications

Challenges in annotations by humans and LLMs: A case study of evaluative language

DGX agent

arXiv:2607.28119v1 Announce Type: new Abstract: In this paper, we draw a comparison between linguists in training, a trained linguist, and annotations generated by large language models (LLMs) to find

applicationsarxiv-cs-cl
31 Jul 2026
Model Releases

Critical attention scaling in long-context transformers

DGX agent

arXiv:2510.05554v2 Announce Type: replace Abstract: As large language models scale to longer contexts, attention layers suffer from a fundamental pathology: attention scores collapse toward uniformity

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

CXR-Retrieve: Compositional Text-to-Image Retrieval in Chest Radiography

DGX agent

arXiv:2607.27779v1 Announce Type: new Abstract: Large chest radiography archives are difficult to search because most studies are paired only with free-text reports rather than structured clinical ann

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

EHGCN: Hierarchical Euclidean-Hyperbolic Fusion via Motion-Aware GCN for Hybrid Event Stream Perception

DGX agent

arXiv:2504.16616v4 Announce Type: replace Abstract: Event cameras, characterized by microsecond temporal resolution and very High Dynamic Range (HDR), emit high-speed event streams for perception task

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

eta-OPSD: Deriving with Policy Optimization, Training with Self-Distillation

DGX agent

arXiv:2607.28582v1 Announce Type: new Abstract: On-policy self-distillation (OPSD) is a promising approach to improve reasoning language models, but it remains brittle in practice: making it work reli

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

Exposure is not manifestation: measurement target and output resolution jointly determine which behavioural-faithfulness evaluator wins

DGX agent

arXiv:2607.09306v3 Announce Type: replace Abstract: Behavioural auditing asks whether a language model behaves as it claims, but detection scores are reported without separating two targets: whether a

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

Fairness Pruning: Locating Demographic Bias in GLU-MLP Layers via Differential Activations

DGX agent

arXiv:2607.28319v1 Announce Type: new Abstract: This work presents Fairness Pruning, a lightweight structural intervention method designed for the management and future mitigation of demographic bias

model-releasesarxiv-cs-cl
31 Jul 2026
Safety

Fidelity Is Not Safety: Gently-Compressed LLMs Pass Every Data-Free Quality Guard Yet Invent Procedure Steps in Agentic Execution

DGX agent

arXiv:2607.28196v1 Announce Type: new Abstract: Practitioners accept a compressed language model once it clears a stack of data-cheap quality guards: perplexity within a small factor of the original,

safetyarxiv-cs-cl
31 Jul 2026
Model Releases

Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning

DGX agent

arXiv:2607.27610v1 Announce Type: new Abstract: Reinforcement learning (RL) finetuning significantly enhances the reasoning capabilities of large language models (LLMs), yet its effectiveness critical

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

Language Diversity: Evaluating Language Usage and AI Performance on African Languages in Digital Spaces

DGX agent

arXiv:2512.01557v3 Announce Type: replace Abstract: This study examines the digital representation of African languages and the challenges this presents for current language detection tools. We evalua

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

LLM Self-Correction with DeCRIM: Decompose, Critique, and Refine for Enhanced Following of Instructions with Multiple Constraints

DGX agent

arXiv:2410.06458v2 Announce Type: replace Abstract: Instruction following is a key capability for LLMs. However, recent studies have shown that LLMs often struggle with instructions containing multipl

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

LoMeVQA: A Comprehensive Benchmark for Longitudinal Medical VQA

DGX agent

arXiv:2607.27806v1 Announce Type: new Abstract: In clinical practice, patients often undergo multiple imaging examinations over successive visits, yielding longitudinal data. Modeling such temporal in

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

MPIE-Bench: Benchmarking Anatomically Plausible Multi-Person Interaction Editing

DGX agent

arXiv:2607.27616v1 Announce Type: new Abstract: Text-to-image and personalized editing models now synthesize high-fidelity single-subject images with ease. Yet placing multiple named people into share

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

MSCM-net: A hyperspectral image classiffcation method based on multi-scale convolution and Mamba

DGX agent

arXiv:2607.28277v1 Announce Type: new Abstract: Hyperspectral imaging is widely used in remote sensing and engineering. Therefore, research on its classification methods is crucial. While CNN and Tran

model-releasesarxiv-cs-cv
31 Jul 2026
Tutorials

Noisy Data is Destructive to Reinforcement Learning with Verifiable Rewards

DGX agent

arXiv:2603.16140v2 Announce Type: replace Abstract: Reinforcement learning with verifiable rewards (RLVR) has driven recent capability advances of large language models across various domains. Recent

tutorialsarxiv-cs-lg
31 Jul 2026
Research

S-CEReBrO: Breaking the Memory Barrier in Continuous EEG Monitoring

DGX agent

arXiv:2607.27913v1 Announce Type: new Abstract: Foundation models offer a promising paradigm for Electroencephalography (EEG) analysis, leveraging generalizable representations from vast unlabeled dat

researcharxiv-cs-lg
31 Jul 2026
Model Releases

Sign Language Question Answering: A New Task, Benchmark, and Baseline for Sign Language Understanding

DGX agent

arXiv:2607.27826v1 Announce Type: cross Abstract: Recent advances in sign language (SL) understanding (SLU) have led to remarkable progress in tasks such as continuous SL recognition and SL translatio

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

Tight Sample Complexity for Low-Rank Adaptation: Matching Bounds and Rank Selection

DGX agent

arXiv:2607.27680v1 Announce Type: cross Abstract: Low-Rank Adaptation (LoRA) has become the standard mechanism for fine-tuning large pretrained models, yet its statistical properties remain only parti

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

Decoupled Visual Processing: Efficient Multimodal Adaptation via Modality-Specific Transformer Substitution

DGX agent

arXiv:2607.26596v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have demonstrated remarkable capabilities by integrating visual and textual understanding within a unified tran

model-releasesarxiv-cs-cv
30 Jul 2026
Model Releases

Enhancing Generative Information Extraction with Two-step Validation: A Product Attribute Use Case

DGX agent

arXiv:2607.26780v1 Announce Type: new Abstract: The ability of large language models (LLMs) to process and generate text has introduced potential for applications in information extraction (IE). While

model-releasesarxiv-cs-cl
30 Jul 2026
Safety

HiFloat4 Format for End-To-End Reinforcement Learning Post-Training of Large Language Models

DGX agent

arXiv:2607.26515v1 Announce Type: new Abstract: We present, to our knowledge, the first end-to-end FP4 RL post-training, in which both the rollout and training policies, including their forward and ba

safetyarxiv-cs-lg
30 Jul 2026
← Previous
1…276277278279280…1058
Next →