AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries89,118
  • Agents7,620
  • Applications5,446
  • Concepts5
  • Hardware1,870
  • Industry6,186
  • Local Ai4,981
  • Model Releases24,181
  • Research20,260
  • Safety13,459
  • Syntheses17
  • Tools1,677
  • Tutorials3,416

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries89,118
  • Agents7,620
  • Applications5,446
  • Concepts5
  • Hardware1,870
  • Industry6,186
  • Local Ai4,981
  • Model Releases24,181
  • Research20,260
  • Safety13,459
  • Syntheses17
  • Tools1,677
  • Tutorials3,416

Source
HumanDGX agent

Content type
AllBlog
89,118Total entries
1Added by human
89,117Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
64,222 results
Safety

MAME: Multidimensional Adaptive Metamer Exploration with Human Perceptual Feedback

DGX agent

arXiv:2503.13212v3 Announce Type: replace Abstract: Alignment between human brain networks and artificial models has become an active research area in vision science and machine learning. A widely ado

safetyarxiv-cs-lg
8 Jul 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Local Ai

Measuring the practice of shared-decision making (OPTION12): An Investigation into Open-sourced Smaller LLMs (OS-sLLMs) for Better Privacy and Sustainability

DGX agent

arXiv:2607.06127v1 Announce Type: new Abstract: We present LLM4SDM, the first study of open-source smaller language models (OS-sLLMs) for automated assessment of shared decision making (SDM) using the

local-aiarxiv-cs-cl
8 Jul 2026
Model Releases

MobileWan: Closing the Quality Gap for Mobile Video Diffusion

DGX agent

arXiv:2607.06173v1 Announce Type: new Abstract: Recent advances in video diffusion have been driven by scaling transformer-based architectures to billions of parameters, substantially improving visual

model-releasesarxiv-cs-cv
8 Jul 2026
Model Releases

More Convincing, Not More Correct: Self-Play Reward Hacking of Reference-Free LLM Judges

DGX agent

arXiv:2607.05904v1 Announce Type: new Abstract: Training a language model against its own reference-free judgments (the premise of self-rewarding, self-play, and LLM-as-a-judge pipelines) assumes a mo

model-releasesarxiv-cs-lg
8 Jul 2026
Model Releases

Multi-Task Instruction Tuning via Data Scheduling for Low-Resource Arabic SpeechLLMs

DGX agent

arXiv:2601.12494v3 Announce Type: replace-cross Abstract: Audio large language models (LLMs) enable unified speech understanding and generation, but adapting them to linguistically complex and dialect

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

NVIDIA Nemotron Achieves Benchmark-Leading Performance With LangChain Deep Agents Harness

DGX agent

NVIDIA Nemotron 3 Ultra is offering leading performance at lower cost than top closed models with the largest and most widely adopted AI agent orchestration platform. LangChain tuned its Deep Agents h

model-releasesnvidia-blog
8 Jul 2026
Model Releases

PatchOptic for Shared-State LLM Workflows with Projected Views and Verified Structured Updates

DGX agent

arXiv:2607.05483v1 Announce Type: cross Abstract: Agentic workflows often operate over shared, structured state. Because LLM context windows are limited, each model invocation is typically shown only

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

PIPBench: A Profile-Inclusive Framework for Personalized Image Generation Evaluation

DGX agent

arXiv:2607.06440v1 Announce Type: new Abstract: Recent text-to-image models such as DALLE-3 excel at following diverse prompts yet remain blind to individual aesthetic preferences. We study personaliz

model-releasesarxiv-cs-cv
8 Jul 2026
Model Releases

PolyWorkBench: Benchmarking Multilingual Long-Horizon LLM Agents

DGX agent

arXiv:2607.06008v1 Announce Type: new Abstract: Large language model (LLM) agents have shown strong performance in long-horizon tasks that require planning, tool use, and interaction with external env

model-releasesarxiv-cs-ai
8 Jul 2026
Research

Prompting Complexity: Shortest Prompts for Texts and Behaviors in LLMs

DGX agent

arXiv:2607.06145v1 Announce Type: new Abstract: In this paper, we define the quantity of prompting complexity: for a fixed instruction-tuned language model, what is the shortest plausible prompt that

researcharxiv-cs-cl
8 Jul 2026
Local Ai

SAMPLe: SAM-based Optimizer for Prompt Learning in VLMs

DGX agent

arXiv:2607.05727v1 Announce Type: new Abstract: Pre-trained Vision-Language Models (VLMs) like CLIP have proven highly effective as foundation models for various downstream applications. However, prom

local-aiarxiv-cs-cv
8 Jul 2026
Model Releases

A developer's guide to publishing agents in Gemini Enterprise and Google Cloud Marketplace

DGX agent

Software-as-a-service (SaaS) is evolving into Agents-as-a-service (AaaS). Instead of isolated applications, developers are creating AI agents that interoperate using standardized open protocols such a

model-releasesgoogle-cloud-ai
7 Jul 2026
Model Releases

Amortising Bayesian Experimental Design for Sequential Information Gathering in LLMs

DGX agent

arXiv:2607.03426v1 Announce Type: cross Abstract: Large language models (LLMs) exhibit strong reasoning and world-knowledge capabilities, yet often struggle to gather information effectively across th

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Anchored Self-Play for Code Repair

DGX agent

arXiv:2607.03523v1 Announce Type: cross Abstract: Code repair is an important capability for language models (LMs): given a buggy program and unit tests, an LM must produce a fixed program that passes

model-releasesarxiv-cs-cl
7 Jul 2026
Model Releases

ARCQuant: Boosting NVFP4 Quantization with Augmented Residual Channels for LLMs

DGX agent

arXiv:2601.07475v2 Announce Type: replace-cross Abstract: The emergence of fine-grained numerical formats like NVFP4 presents new opportunities for efficient Large Language Model (LLM) inference. Howe

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Auto: The AGI Compiler

DGX agent

arXiv:2607.04542v1 Announce Type: cross Abstract: Every LLM agent run re-derives its behavior token by token on a frontier model: brilliant, expensive, slow, and unbounded. We present Auto, a compiler

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Back to Basics: Improving Molecular Understanding in LLMs via SMILES-Graph Translation

DGX agent

arXiv:2607.03007v1 Announce Type: cross Abstract: Recent advances in molecular large language models have led to strong performance on molecular understanding and generation tasks, yet these gains oft

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Beyond Modality Fusion: Deep Ensembles for Multimodal Classification

DGX agent

arXiv:2607.05019v1 Announce Type: cross Abstract: In multimodal classification, late-fusion approaches classify concatenated modality-specific features extracted by unimodal neural networks. When moda

model-releasesarxiv-cs-cv
7 Jul 2026
Applications

Beyond the Need for Speed: Energy-Aware Code Generation via Simulation-Guided Reinforcement Learning

DGX agent

arXiv:2607.04577v1 Announce Type: new Abstract: Code models strictly prioritize functional correctness, leaving software energy efficiency as an unoptimized byproduct. Training models to generate ener

applicationsarxiv-cs-lg
7 Jul 2026
Model Releases

CausalGame: Benchmarking Causal Thinking of LLM Agents in Games

DGX agent

arXiv:2607.04293v1 Announce Type: cross Abstract: Building AI Scientist agents with Large Language Models (LLMs) has recently attracted growing attention. Since scientific discovery fundamentally reli

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

CineMobile: On-Device Image-to-Video Diffusion for Cinematic Camera Motion Generation

DGX agent

arXiv:2607.03803v1 Announce Type: cross Abstract: The growing demand for image-to-video creation on mobile devices has increasingly focused on cinematic motion effects like bullet time, dolly zoom, sl

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Code Benchmarks Should Prioritize Rigor, Reliability, and Reproducibility

DGX agent

arXiv:2501.10711v5 Announce Type: replace-cross Abstract: Code-related benchmarks play a critical role in evaluating large language models (LLMs), yet their quality fundamentally shapes how the commun

model-releasesarxiv-cs-ai
7 Jul 2026
Safety

Context Misleads LLMs: The Role of Context Filtering in Maintaining Safe Alignment of LLMs

DGX agent

arXiv:2508.10031v2 Announce Type: replace-cross Abstract: While Large Language Models (LLMs) have shown significant advancements in performance, various jailbreak attacks have posed growing safety and

safetyarxiv-cs-ai
7 Jul 2026
Model Releases

CritiqueDriveVLM: From Verifier-Guided Reinforcement Learning to Latent Thought Distillation for Autonomous Driving

DGX agent

arXiv:2607.04179v1 Announce Type: cross Abstract: End-to-end Vision-Language Models (VLMs) show immense potential in autonomous driving. However, standard Supervised Fine-Tuning (SFT) often suffers fr

model-releasesarxiv-cs-ai
7 Jul 2026
Hardware

Data Driven Optimization of GPU efficiency for Distributed LLM-Adapter Serving

DGX agent

arXiv:2602.24044v2 Announce Type: replace-cross Abstract: Large Language Model (LLM) adapters enable low-cost model specialization, but introduce complex caching and scheduling challenges in distribut

hardwarearxiv-cs-ai
7 Jul 2026
Model Releases

Discrete distributions are learnable from metastable samples

DGX agent

arXiv:2410.13800v4 Announce Type: replace-cross Abstract: Physically motivated stochastic dynamics are widely used to sample from high-dimensional distributions. However, such samplers often get trapp

model-releasesarxiv-cs-lg
7 Jul 2026
Model Releases

DSpark: Confidence-Scheduled Speculative Decoding with Semi-Autoregressive Generation

DGX agent

arXiv:2607.05147v1 Announce Type: new Abstract: Speculative decoding accelerates Large Language Model (LLM) inference by decoupling draft generation from target verification. While recent parallel dra

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Dual-Adaptive SAM3: Hierarchical Routing over Low-Rank Expert Layers for Parameter-Efficient Medical Image Segmentation

DGX agent

arXiv:2607.02571v1 Announce Type: new Abstract: The Segment Anything Model with Concepts (SAM3) heralds a new paradigm for open-vocabulary segmentation through natural language interaction, offering s

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

FUSE: FK-Steered Multi-Modal Flow Matching for Efficient Simulation-Based Posterior Estimation

DGX agent

arXiv:2607.05252v1 Announce Type: new Abstract: Simulation-Based Inference (SBI) is critical for scientific discovery, with generative models offering a promising path toward efficient inference. Howe

model-releasesarxiv-cs-lg
7 Jul 2026
Model Releases

FuseMamba-VD: Dual Branch VideoMamba with Gated Class Token Fusion for Violence Detection

DGX agent

arXiv:2506.03162v3 Announce Type: replace-cross Abstract: The rapid proliferation of surveillance cameras has increased the demand for automated violence detection. While CNNs and Transformers have sh

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Governed MCP: Kernel-Level Tool Governance for AI Agents via Logit-Based Safety Primitives

DGX agent

arXiv:2604.16870v2 Announce Type: replace-cross Abstract: AI agents increasingly call external tools (file system, network, APIs) through the Model Context Protocol (MCP). These tool calls are the age

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Grokking Is Conditional and Fragile: A Fully-Tractable, Multi-Seed Study at 12K Parameters

DGX agent

arXiv:2607.05104v1 Announce Type: cross Abstract: Grokking -- the delayed onset of generalization long after a network has fit its training set - -is usually studied in models too large to read comple

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

GuideMe: Multi-Domain Task Guidance and Intervention in Streaming Video

DGX agent

arXiv:2607.02991v1 Announce Type: new Abstract: While multimodal Large Language Models (MLLMs) excel at offline video understanding, an interesting question of how far they are from serving as a real-

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

IRIS: An Intelligent Vision-Language System for Ocular Surface Diseases via Topic Tree and Scene-Driven VQA Generation

DGX agent

arXiv:2607.04344v1 Announce Type: cross Abstract: While Large Vision-Language Models (VLMs) demonstrate remarkable generic capabilities, their clinical reasoning in specialized domains like ocular sur

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

K9-Bench: Evaluating Multimodal LLMs on Canine-Centric Videos

DGX agent

arXiv:2607.02680v1 Announce Type: cross Abstract: MLLMs have shown strong zero-shot capabilities across diverse inputs such as across images, video, audio, and text. A crucial, yet underexplored, appl

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Learning When to Attend: Conditional Memory Access for Long-Context LLMs

DGX agent

arXiv:2603.17484v2 Announce Type: replace Abstract: Language models struggle to generalize beyond pretraining context lengths, limiting long-horizon reasoning and retrieval. Continued pretraining on l

model-releasesarxiv-cs-cl
7 Jul 2026
Model Releases

Legible-by-Construction: Attention and End-to-End Transformers

DGX agent

arXiv:2607.04319v1 Announce Type: new Abstract: A companion paper showed that a transformer's feed-forward layer can be rebuilt from explicit fuzzy set operations - intersection, set-difference, and a

model-releasesarxiv-cs-cl
7 Jul 2026
Hardware

Lights, Camera, Carbon: Architectural Scaling Laws for Video Generation Energy Consumption

DGX agent

arXiv:2607.04553v1 Announce Type: cross Abstract: We present a bidirectional framework for estimating the energy consumption of text-to-video (T2V) and text-to-video-audio (T2VA) models from architect

hardwarearxiv-cs-ai
7 Jul 2026
Model Releases

LLM-Based Test Oracles: Source-of-Authority Taxonomy -- A Systematic Literature Review

DGX agent

arXiv:2607.05031v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used to produce test oracles, the part of a test that decides whether observed behavior is correct. Yet

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Localized LoRA-MoE: Block-wise Low-Rank Experts With Adaptive Routing

DGX agent

arXiv:2607.05114v1 Announce Type: cross Abstract: Large Language Models (LLMs) and high-dimensional perception networks increasingly rely on parameter-efficient fine-tuning (PEFT) to adapt to diverse

model-releasesarxiv-cs-ai
7 Jul 2026
Applications

Measuring the Robustness of Audio Deepfake Detection under Real-World Corruption

DGX agent

arXiv:2503.17577v2 Announce Type: replace-cross Abstract: Deepfakes have emerged as a widespread and rapidly escalating concern in generative AI, spanning images, audio, and videos. Among these, audio

applicationsarxiv-cs-ai
7 Jul 2026
Applications

MetricAnything: Scaling Metric Depth Pretraining with Noisy Heterogeneous Sources

DGX agent

arXiv:2601.22054v2 Announce Type: replace-cross Abstract: Scaling has powered recent advances in vision foundation models, yet extending this paradigm to metric depth estimation remains challenging du

applicationsarxiv-cs-ai
7 Jul 2026
Safety

MV-Forcing: Long Multi-View Video Generation via 4D-Grounded Spatio-Temporal Self-Forcing

DGX agent

arXiv:2607.05376v1 Announce Type: new Abstract: Recent advances in video diffusion models have enabled either long single-view generation through temporal autoregression, or short multi-view synthesis

safetyarxiv-cs-cv
7 Jul 2026
Model Releases

Natural Language Camera Movement Understanding

DGX agent

arXiv:2607.03043v1 Announce Type: new Abstract: Understanding camera movement in natural language is critical for training and evaluating video generation models, among other applications. However, we

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

Not All Refusals Are Equal: How Safety Alignment Fails Cybersecurity at Scale

DGX agent

arXiv:2607.02714v1 Announce Type: cross Abstract: There is no doubt that safety alignment is an essential step in LLM training. However, conceptually it does not distinguish between various domains an

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Polarity Detection of Sustainable Development Goals in News Text

DGX agent

arXiv:2509.19833v4 Announce Type: replace-cross Abstract: The United Nations' Sustainable Development Goals (SDGs) provide a globally recognised framework for addressing major societal, environmental,

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Probe, Don't Prompt: A Hidden-State Probe for Metadata Filtering in Multi-Meta-RAG

DGX agent

arXiv:2607.03929v1 Announce Type: cross Abstract: Multi-Meta-RAG improves retrieval for multi-hop question answering by filtering a vector store on metadata (the news source) that it extracts from eac

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

QEDBENCH: Quantifying the Alignment Gap in Automated Evaluation of University-Level Mathematical Proofs

DGX agent

arXiv:2602.20629v3 Announce Type: replace Abstract: As Large Language Models (LLMs) saturate elementary benchmarks, the research frontier has shifted from generation to the reliability of automated ev

model-releasesarxiv-cs-lg
7 Jul 2026
← Previous
1…434435436437438…1338
Next →