AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
HumanDGX agent

Content type
87,814Total entries
1Added by human
87,813Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,587 results
Model Releases

Beyond Liars' Bench: The Impact of Lie Typology, Depth, and Sparsity on Deception Detection in LLMs

DGX agent

arXiv:2607.20479v1 Announce Type: new Abstract: Training probes to detect deceptive outputs from large language models is still an open problem. Recent work has demonstrated that detection probes fail

model-releasesarxiv-cs-ai
24 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Beyond Sycophancy: Structured Resistance and Compliance in LLM Moral Reasoning

DGX agent

arXiv:2607.21558v1 Announce Type: new Abstract: Building socially calibrated large language models, which can learn from others without simply yielding to them, requires more than reducing sycophancy

safetyarxiv-cs-ai
24 Jul 2026
Model Releases

CAMeR: Keyword-Gated Hybrid Activation for Adaptive Memory Retention in LLM Agents

DGX agent

arXiv:2607.20458v1 Announce Type: cross Abstract: Large language model (LLM) agents operating over extended dialogues accumulate vast amounts of information, yet existing memory systems either retain

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

CRAG-MM-Diagnostics: Enabling Stage-Wise Analysis of Knowledge-Intensive VQA

DGX agent

arXiv:2607.21155v1 Announce Type: cross Abstract: Knowledge-Intensive Visual Question Answering (KI-VQA) benchmarks evaluate Vision-Language Models (VLMs) as multimodal knowledge assistants by requiri

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

DFAH-Bench: Benchmarking Observable Agent Instability in Financial Decision-Making

DGX agent

arXiv:2607.20491v1 Announce Type: new Abstract: Standard evaluation benchmarks measure what a tool-using agent decides, not whether it arrives at that decision through the same process each time. We i

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Dropping the Anchor: Statistical Context Summarization for Distributed Systems via Pulsar Attention

DGX agent

arXiv:2607.20457v1 Announce Type: cross Abstract: Inference with large language models (LLMs) on long sequences is computationally expensive due to the quadratic complexity of self-attention. Distribu

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Evaluating the Effectiveness of Persona Simulation in Opinion Prediction with GPT-4.1

DGX agent

arXiv:2607.20589v1 Announce Type: new Abstract: Persona simulation involves utilizing large language models (LLMs) to anticipate human choices or interactions based on specific characteristic informat

model-releasesarxiv-cs-cl
24 Jul 2026
Model Releases

GaugeQuant: Online Learning of Quantization-Optimal Bases from LLM Symmetries

DGX agent

arXiv:2607.20757v1 Announce Type: cross Abstract: Transformers are known to have internal continuous symmetries that leave outputs invariant, while modifying quantization. GaugeQuant leverages this in

model-releasesarxiv-cs-cl
24 Jul 2026
Model Releases

Geometric Configurations of Perturbed Jailbreak Prompts

DGX agent

arXiv:2607.20581v1 Announce Type: cross Abstract: Perturbation techniques that turn unsuccessful jailbreak prompts into successful ones are continuously evolving, constituting a major security threat

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Gradient Concentration, Not Weight Saliency, Explains Representation-Level Class Unlearning

DGX agent

arXiv:2607.21353v1 Announce Type: new Abstract: Machine unlearning aims to remove the influence of specific training data while preserving model utility. Many state-of-the-art approaches pursue this g

model-releasesarxiv-cs-lg
24 Jul 2026
Model Releases

Is MoE Routing a Huffman Code? Discovering the Frequency-Diversity Law in Chain-of-Thought

DGX agent

arXiv:2607.20427v1 Announce Type: cross Abstract: Mixture-of-Experts architectures have revolutionized scaling, yet the underlying logic of their routing remains a black box. In this paper, we uncover

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Isolating LLM Alignment from Regex: Zero Coverage and Metric-Dependent Divergence Under Adversarial Mutation

DGX agent

arXiv:2607.20494v1 Announce Type: new Abstract: Production LLM applications commonly stack a regex filter in front of model-side alignment; prior work found no measurable coverage gain from adding a l

model-releasesarxiv-cs-ai
24 Jul 2026
Research

Marking the Wrong Symptoms: Evaluating LLM Watermarks in Medical Texts

DGX agent

arXiv:2607.20462v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly integrated into clinical workflows, stressing the need for reliable traceability of model-generated output

researcharxiv-cs-ai
24 Jul 2026
Hardware

NVIDIA-labs OO Agents: Native Python Object-Oriented Agents

DGX agent

arXiv:2607.20709v1 Announce Type: new Abstract: Traditional agent development is split across prompt templates, tool schemas, callback code, and workflow graphs. We present NVIDIA Object-Oriented Agen

hardwarearxiv-cs-ai
24 Jul 2026
Model Releases

PersonaTrail: Benchmarking Personalized Web Agents through Browsing Trails

DGX agent

arXiv:2607.20482v1 Announce Type: new Abstract: Recent advances in large language models have enabled web agents to autonomously execute complex tasks. In practice, users frequently provide underspeci

model-releasesarxiv-cs-ai
24 Jul 2026
Safety

PILD: Physics-Informed Learning via Diffusion

DGX agent

arXiv:2601.21284v2 Announce Type: replace-cross Abstract: Diffusion models have emerged as powerful generative tools for modeling complex data distributions, yet their purely data-driven nature limits

safetyarxiv-cs-ai
24 Jul 2026
Model Releases

PromptPack: Scaling LLM Annotation Agents for Online Recommendation

DGX agent

arXiv:2607.20528v1 Announce Type: new Abstract: Online recommendation platforms increasingly use Large Language Models (LLMs) to extract structured features from ad creatives. While deploying a single

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

RE-AD: Real-Time Requirement Adherence for Data Labeling

DGX agent

arXiv:2607.20455v1 Announce Type: cross Abstract: Human-annotated data remains fundamental to training frontier Large Language Models (LLMs). However, crowd-sourced annotations often suffer from quali

model-releasesarxiv-cs-ai
24 Jul 2026
Safety

Robust Critics: Defending LLMs Against Multi-Turn Attacks

DGX agent

arXiv:2607.20472v1 Announce Type: new Abstract: When a user asks a language model something harmful, is it a genuine attack or a misunderstood but well-meaning question? This ambiguity is one of the c

safetyarxiv-cs-ai
24 Jul 2026
Local Ai

Routing Without Training: Controllable-Ratio LLM Offloading via Reliability Gating

DGX agent

arXiv:2607.20481v1 Announce Type: new Abstract: Local-cloud collaboration is a practical way to deploy large language models under resource constraints, but existing methods often rely on trained rout

local-aiarxiv-cs-ai
24 Jul 2026
Model Releases

Scene Parameter Saliency via Differentiable Light Transport

DGX agent

arXiv:2607.21562v1 Announce Type: new Abstract: Gradient-based saliency methods reveal which input features most influence a neural network's output, and are a standard tool for model interpretability

model-releasesarxiv-cs-cv
24 Jul 2026
Tutorials

SCoPE: Shift-Aware Speaker-Conditioned Priors for Emotion Recognition in Conversations

DGX agent

arXiv:2607.20445v1 Announce Type: new Abstract: In conversations, human emotions are transient; however, they tend to persist across multiple utterances. For example, we rarely switch instantly betwee

tutorialsarxiv-cs-cl
24 Jul 2026
Applications

SoccerSynth Field: enhancing field detection with synthetic data from virtual soccer simulator

DGX agent

arXiv:2503.13969v2 Announce Type: replace Abstract: Field detection in team sports is an essential task in sports video analysis. However, collecting large-scale and diverse real-world datasets for tr

applicationsarxiv-cs-cv
24 Jul 2026
Model Releases

StabilityBench: Benchmarking Instability in LLMs

DGX agent

arXiv:2607.20558v1 Announce Type: cross Abstract: AI Assistants are increasingly deployed in high-stakes settings, such as healthcare or government services. Yet their real-world behavior remains poor

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Student-Centered Distillation Narrows the Agentic Gap Between Small and Large LLMs

DGX agent

arXiv:2509.14257v3 Announce Type: replace-cross Abstract: Large Language Model agents achieve strong performance on multi-step reasoning and tool-use tasks, but their impressive capabilities typically

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Tencent WorkBuddy Bench: A Multi-Domain Coding-Agent Benchmark with Contamination-Resistant Task Construction

DGX agent

arXiv:2607.20911v1 Announce Type: new Abstract: We introduce Tencent WorkBuddy Bench, a multi-domain evaluation suite for coding agents; this report documents its construction methodology, scoring pro

model-releasesarxiv-cs-cl
24 Jul 2026
Model Releases

The Devil is in the Spectrum: Mitigating Representation Collapse in LLMs via Topologically Regularized Side-Path

DGX agent

arXiv:2607.20484v1 Announce Type: new Abstract: Large Language Models (LLMs) are fundamentally limited by representation collapse, a bottleneck that severely degrades long-context performance. We iden

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

The Geometry of Personality: Activation Steering with Jungian Cognitive Functions

DGX agent

arXiv:2607.20803v1 Announce Type: cross Abstract: Activation steering enables control and interpretation of LLMs, yet existing work primarily models personality through static trait frameworks such as

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Three-Pronged Spectral Control for Federated Parameter Efficient Fine Tuning

DGX agent

arXiv:2607.20914v1 Announce Type: new Abstract: Federated parameter-efficient fine-tuning (PEFT) enables communication-efficient adaptation of large pretrained models on decentralized edge data, but i

model-releasesarxiv-cs-lg
24 Jul 2026
Model Releases

WebCoach: Self-Evolving Web Agents with Cross-Session Memory Guidance

DGX agent

arXiv:2511.12997v2 Announce Type: replace Abstract: Multimodal LLM-powered agents have recently demonstrated impressive capabilities in web navigation, enabling agents to complete complex browsing tas

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Webly Supervised Multi-Label Recognition: Evaluation Benchmark and Dual-Branch Multi-Label Contrastive Learning

DGX agent

arXiv:2607.20874v1 Announce Type: new Abstract: Training deep learning models with freely available web images can reduce their dependence on costly manual annotations. Although webly supervised learn

model-releasesarxiv-cs-cv
24 Jul 2026
Model Releases

What is Good? Extracting and Testing Implicit Theories of Literary Quality from LLM Reasoning Traces

DGX agent

arXiv:2607.20425v1 Announce Type: new Abstract: What makes writing 'good' remains a persistent question in literary studies and computational linguistics. We present a two-study investigation of how r

model-releasesarxiv-cs-cl
24 Jul 2026
Model Releases

When Are Reasoning-Based Guardrails Not Efficient? ResponseGuard: A Fast Vision-Language Guard for Real-Time Moderation

DGX agent

arXiv:2607.21401v1 Announce Type: cross Abstract: A vision-language AI assistant returns its answer as a stream of generated tokens. Therefore, a safety guard that watches that answer has to keep up w

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Windowed-MTP: Removing the Full-Context Draft-KV Tax at Million-Token Context

DGX agent

arXiv:2607.21535v1 Announce Type: cross Abstract: Speculative decoding accelerates autoregressive generation by having a cheap draft propose tokens that a target verifies in parallel. Frontier models

model-releasesarxiv-cs-cl
24 Jul 2026
Model Releases

Workflow-Localized Mechanism Learning: Attribution-Guided Repair and Knowledge Reuse for Structured Agent Skills

DGX agent

arXiv:2607.20999v1 Announce Type: new Abstract: Agent Skills package reusable procedural knowledge as external artifacts for frozen language-model agents, yet existing optimizers do not jointly resolv

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

A Novel Hybrid Deep Learning Technique for Speech Emotion Detection using Feature Engineering

DGX agent

arXiv:2507.07046v3 Announce Type: replace-cross Abstract: Nowadays, speech emotion recognition (SER) plays a vital role in the field of human-computer interaction (HCI) and the evolution of artificial

model-releasesarxiv-cs-ai
23 Jul 2026
Agents

Agent-Centric Animal Pose Forecasting

DGX agent

arXiv:2607.19548v1 Announce Type: new Abstract: Understanding animal behavior at an algorithmic level -- what animals attend to, how they form internal models and plans, and how this maps to action --

agentsarxiv-cs-lg
23 Jul 2026
Model Releases

Alipay-PIBench: A Realistic Payment Integration Benchmark for Coding Agents

DGX agent

arXiv:2607.14573v3 Announce Type: replace Abstract: Payment integration is a demanding repository-level software task: agents must select a suitable product, implement coordinated client-server flows,

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

BaseRT: Advancing Best-in-Class LLM Inference with Apple M5 Neural Accelerators

DGX agent

arXiv:2607.19438v1 Announce Type: cross Abstract: Apple's M5 generation introduces a redesigned GPU architecture in which every core carries a dedicated Neural Accelerator: on-die matrix units exposed

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

Beyond Relevance-Centric Retrieval: Rubric-Oriented Document Set Selection and Ranking

DGX agent

arXiv:2607.19747v1 Announce Type: new Abstract: As large language models and AI agents become the primary consumers of search results, document set quality determines the upper bound of downstream gen

model-releasesarxiv-cs-cl
23 Jul 2026
Model Releases

CreatiPoster: Towards Editable and Controllable Multi-Layer Graphic Design Generation

DGX agent

arXiv:2506.10890v2 Announce Type: replace Abstract: Graphic design plays a crucial role in both commercial and personal contexts, yet creating high-quality, editable, and aesthetically pleasing graphi

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

DobicVLM: Aligning Chest X-Ray Report Generation with Clinically-Grounded Programmatic Rewards via Group Relative Policy Optimization

DGX agent

arXiv:2607.18988v1 Announce Type: new Abstract: Medical imaging is a cornerstone of diagnostics, yet automated chest X-ray report generation struggles with structural adherence, anatomical completenes

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

FORCE-Bench: A Benchmark, Dataset, and Evaluation Harness for Agentic AI in Enterprise Finance

DGX agent

arXiv:2607.19409v1 Announce Type: new Abstract: Recent advances in large language models have accelerated deployment of agentic systems in operational finance. Existing benchmarks emphasize measuring

model-releasesarxiv-cs-ai
23 Jul 2026
Research

Generative World Renderer at the Speed of Play

DGX agent

arXiv:2607.18703v1 Announce Type: new Abstract: Generative world renderer AlayaRenderer receives structured world states exported from physics engines and synthesizes RGB frames. Unlike models that ge

researcharxiv-cs-cv
23 Jul 2026
Model Releases

GLID: Gated Local Intrinsic Dimension Repairs the Blind Spots of Face-Forgery Detectors

DGX agent

arXiv:2607.18770v1 Announce Type: cross Abstract: Fine-tuned foundation-model detectors dominate face-forgery benchmarks, yet they stay blind to generator families absent from training. We present GLI

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

IBoxCLA: Towards Robust Box-supervised Segmentation of Polyp via Improved Box-dice and Contrastive Latent-anchors

DGX agent

arXiv:2310.07248v5 Announce Type: replace Abstract: Box-supervised polyp segmentation attracts increasing attention for its cost-effective potential. Existing solutions often rely on learning-free met

model-releasesarxiv-cs-cv
23 Jul 2026
Safety

LatentLens: Revealing Highly Interpretable Visual Tokens in LLMs

DGX agent

arXiv:2602.00462v5 Announce Type: replace-cross Abstract: Transforming a large language model (LLM) into a vision-language model (VLM) can be achieved by mapping the visual tokens from a vision encode

safetyarxiv-cs-ai
23 Jul 2026
Local Ai

LAVIFT: Latent-Action-Guided Vision Fine-Tuning for Surgical Interaction Recognition

DGX agent

arXiv:2607.19889v1 Announce Type: new Abstract: Understanding instrument-tissue interactions is essential for context-aware surgical AI and autonomous robotic surgery. Pretrained vision-language model

local-aiarxiv-cs-cv
23 Jul 2026
← Previous
1…399400401402403…1075
Next →