AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,457
  • Agents7,399
  • Applications5,302
  • Concepts5
  • Hardware1,786
  • Industry6,117
  • Local Ai4,835
  • Model Releases23,193
  • Research19,715
  • Safety13,094
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,457
  • Agents7,399
  • Applications5,302
  • Concepts5
  • Hardware1,786
  • Industry6,117
  • Local Ai4,835
  • Model Releases23,193
  • Research19,715
  • Safety13,094
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
86,457Total entries
1Added by human
86,456Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
50,764 results
Research

Why the Third Axis Is Freedom

DGX agent

arXiv:2608.05423v1 Announce Type: cross Abstract: In generative training, a model produces an output and is penalised for its difference from an example. With one output per comparison, a model that p

researcharxiv-cs-ai
7 Aug 2026
Research

Attention, Anomalies! Handling Attention Layers in Unsupervised Federated Outlier Detection

DGX agent
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2608.04753v1 Announce Type: new Abstract: Attention layers are the backbone of today's most powerful and impactful models. Models with multi-million and billion parameters rely on contextual kno

researcharxiv-cs-lg
6 Aug 2026
Model Releases

Can Post-Training Transform LLMs into Causal Reasoners?

DGX agent

arXiv:2602.06337v2 Announce Type: replace-cross Abstract: Causal inference is essential for decision-making but remains challenging for non-experts. While large language models (LLMs) show promise in

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

CoCo-IR: Contextual Composed Image Retrieval

DGX agent

arXiv:2608.05149v1 Announce Type: new Abstract: Current instruction-based image retrieval systems are powerful but limited to single-turn interactions, failing to capture the iterative nature of compl

model-releasesarxiv-cs-cv
6 Aug 2026
Hardware

ContextMaster: Interactive Multi-Shot Video Creation via Fixed-Budget Sparse Context Routing

DGX agent

arXiv:2608.04956v1 Announce Type: new Abstract: Recent video models increasingly support generation, reference conditioning, and editing within a single model, yet typically expose them as separate op

hardwarearxiv-cs-cv
6 Aug 2026
Model Releases

Echo Flow Networks

DGX agent

arXiv:2509.24122v3 Announce Type: replace Abstract: At the heart of time-series forecasting (TSF) lies a fundamental challenge: how can models efficiently and effectively capture long-range temporal d

model-releasesarxiv-cs-lg
6 Aug 2026
Model Releases

FinPerMA: A Theory-Informed, Event-Grounded Personalized-Memory Benchmark for LLM Agents

DGX agent

arXiv:2608.04095v1 Announce Type: new Abstract: Large language model (LLM) agents are increasingly used as personalized assistants in high-stakes domains such as financial advising, yet it remains unc

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

FinReportBench: Measuring and Improving Institution-Grade Financial Report Generation

DGX agent

arXiv:2608.04374v1 Announce Type: cross Abstract: Large language models can produce fluent financial analysis, but fluency alone does not establish whether a report is suitable for institutional deliv

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

K-EXAONE 2.0 Technical Report

DGX agent

arXiv:2608.04505v1 Announce Type: new Abstract: This technical report presents K-EXAONE 2.0, an open-weight multilingual foundation model developed by LG AI Research as a step in our effort toward glo

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

Kathleen Writes: Autoregressive Generation and Data Scaling Without Attention

DGX agent

arXiv:2608.04678v1 Announce Type: new Abstract: Papers 1-2 of the Kathleen series showed that a byte-level, attention-free architecture built from a wavetable encoder and multi-scale reverberant state

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

OmniEdit-Bench: A Comprehensive Benchmark for Instruction-based Video Editing

DGX agent

arXiv:2608.05049v1 Announce Type: new Abstract: Instruction-based video editing (IVE) is an emerging field with broad applications, yet evaluating editing models remains challenging. Existing benchmar

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

Reading Between the Frames: Interpreting Implicit and Non-literal Meaning in Social Media Videos

DGX agent

arXiv:2608.04939v1 Announce Type: new Abstract: Social media videos often communicate meanings that go beyond their visible actions, captions, or speech. A mundane clip may become humorous, ironic, or

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

RepoProbe: Benchmarking Architecture-Aware Repository Comprehension with Checklists

DGX agent

arXiv:2608.04783v1 Announce Type: cross Abstract: The integration of Large Language Models (LLMs) into software engineering has shifted the focus from function-level generation to repository-scale ass

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

RooflineBench: A Benchmarking Framework for On-Device LLMs via Roofline Analysis

DGX agent

arXiv:2602.11506v4 Announce Type: replace-cross Abstract: The transition toward localized intelligence through Small Language Models (SLMs) has intensified the need for rigorous performance characteri

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

SEAR: Simple and Efficient Adaptation of Visual Geometric Transformers for Unpaired RGB+Thermal 3D Reconstruction

DGX agent

arXiv:2603.18774v2 Announce Type: replace Abstract: Foundational feed-forward visual geometry models enable accurate and efficient camera pose estimation and scene reconstruction by learning strong sc

model-releasesarxiv-cs-cv
6 Aug 2026
Safety

Social Pressure Breaks Majority Voting in LLM Safety Panels

DGX agent

arXiv:2608.04415v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used to detect unsafe content. A common approach is to combine judgments from a panel of models to correct

safetyarxiv-cs-cl
6 Aug 2026
Model Releases

Teaching Nemotron Greek: Mining a Corpus, Adapting Retrieval, and Grounding Generation for Modern Greek across Specialist Domains

DGX agent

arXiv:2608.05138v1 Announce Type: cross Abstract: Modern Greek is absent from NVIDIA's Nemotron retrieval models and from major multilingual retrieval benchmarks, despite being important for retrieval

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Text2GraphQuery-Bench: A Text to Graph Query Benchmark

DGX agent

arXiv:2602.11745v2 Announce Type: replace Abstract: Graph models are fundamental to data analysis in domains rich with complex relationships. Unlike SQL, which benefits from a rel- atively unified sta

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

When Prompts Become Pixels: Prompt-Region Grounding for Multimodal Reasoning

DGX agent

arXiv:2608.04726v1 Announce Type: new Abstract: Multimodal large language models increasingly reason over screenshots and documents where the task itself may be written in pixels. Yet benchmarks usual

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

AI Security Leaderboard: Methodology, Results and Minimal Standard

DGX agent

arXiv:2608.03070v1 Announce Type: cross Abstract: Frontier AI model developers increasingly rely on layered safeguards to prevent catastrophic misuse, but little public evidence exists on how much pro

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

ChartAnno: Evaluating MLLMs for Chart Annotation Generation

DGX agent

arXiv:2608.03464v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have made significant progress in chart understanding, generation, and editing, but their ability to annotate e

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

DenialRAG: Single-Document RAG Poisoning via Embedded Parametric Denial

DGX agent

arXiv:2608.02678v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) systems are vulnerable to corpus poisoning: an attacker who inserts a crafted document into the retrieval corpus

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

EduClaw-Bench: A Long-Horizon Benchmark for Pedagogical LLM Agents with Simulated Learners

DGX agent

arXiv:2608.03206v1 Announce Type: cross Abstract: Large language models (LLMs) power educational applications from tutoring to essay scoring, but each is a point solution to a single task, and only re

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

FinVerse: Financial Time-Series Benchmark

DGX agent

arXiv:2608.03259v1 Announce Type: cross Abstract: As time-series foundation models have emerged, the need for benchmarks that can evaluate their forecasting ability in meaningful ways has become incre

model-releasesarxiv-cs-ai
5 Aug 2026
Safety

Improved Quantum Algorithms for Reinforcement Learning Under a Generative Model

DGX agent

arXiv:2608.02826v1 Announce Type: cross Abstract: Reinforcement learning is a subfield of machine learning that studies how an agent interacts with an environment in order to extract as large a reward

safetyarxiv-cs-ai
5 Aug 2026
Agents

LACE: Large Language Model Aided Multi-Agent Framework for Agile RISC-V Instruction Extension

DGX agent

arXiv:2608.02915v1 Announce Type: cross Abstract: Domain-specific Instruction Set Architecture eXtensions (ISAX) are widely adopted in the RISC-V ecosystem to accelerate emerging workloads, but implem

agentsarxiv-cs-cl
5 Aug 2026
Applications

MaterialFusion: High-Quality, Zero-Shot, and Controllable Material Transfer with Diffusion Models

DGX agent

arXiv:2502.06606v3 Announce Type: replace Abstract: Manipulating the material appearance of objects in images is critical for applications like augmented reality, virtual prototyping, and digital cont

applicationsarxiv-cs-cv
5 Aug 2026
Research

PASE: Leveraging the Phonological Prior of WavLM for Low-Hallucination Generative Speech Enhancement

DGX agent

arXiv:2511.13300v1 Announce Type: cross Abstract: Generative models have shown remarkable performance in speech enhancement (SE), achieving superior perceptual quality over traditional discriminative

researcharxiv-cs-ai
5 Aug 2026
Model Releases

Route-Align-Verify for Functional Correctness in Code Generation

DGX agent

arXiv:2608.03341v1 Announce Type: cross Abstract: Large language models (LLMs) have substantially improved code generation, yet achieving strong functional correctness remains difficult, especially fo

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Single Canonical Prompts Underestimate LLM Safety's Surface-Form Sensitivity

DGX agent

arXiv:2608.02665v1 Announce Type: cross Abstract: A benchmark score is a measurement instrument, yet most benchmarks read each item at a single canonical surface form. We ask whether that reading is f

model-releasesarxiv-cs-ai
5 Aug 2026
Safety

SPADE: An Input-Adaptive Sparse Attention Engine for Fast Video Diffusion Models Inference

DGX agent

arXiv:2608.03335v1 Announce Type: new Abstract: Video diffusion transformers (vDiTs) generate high quality but pay quadratic self-attention cost, making inference prohibitive at video-token scales. Th

safetyarxiv-cs-cv
5 Aug 2026
Local Ai

The Tell-Tale Trace: Detecting Reasoning Failures in LLMs Using Chain-of-Thought Dynamics

DGX agent

arXiv:2608.03291v1 Announce Type: cross Abstract: Chain-of-thought (CoT) reasoning improves large language model (LLM) performance while also providing an observable interface to the model's reasoning

local-aiarxiv-cs-ai
5 Aug 2026
Research

UHP Detection: LVLMs have their Unique Hallucination Pattern in the Consistency Space

DGX agent

arXiv:2608.03817v1 Announce Type: cross Abstract: Large vision--language models (LVLMs) demonstrate strong multimodal reasoning capabilities but remain prone to hallucination, where model predictions

researcharxiv-cs-ai
5 Aug 2026
Model Releases

Wiring Beats Blending: What Transfers Between Transformer Sizes -- and What Doesn't

DGX agent

arXiv:2608.02829v1 Announce Type: new Abstract: Model families train every size from scratch. Can a pretrained large model be converted into a smaller sibling? We characterize the 1.4B->410M conversio

model-releasesarxiv-cs-lg
5 Aug 2026
Model Releases

A Benchmark Dataset for MLLM-Generated Image Detection: GPT Image2 & Nano Banana2

DGX agent

arXiv:2608.01258v1 Announce Type: new Abstract: The realism of images generated by multimodal large language models (MLLMs), such as GPT Image2 and Nano Banana2, has improved rapidly in recent years.

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

AdaMTP: An Adaptive Training Paradigm for Multi-Token Prediction

DGX agent

arXiv:2608.00434v1 Announce Type: new Abstract: Multi-Token Prediction (MTP) has emerged as an effective paradigm that augments a shared Large Language Model backbone with auxiliary heads, training th

model-releasesarxiv-cs-cl
4 Aug 2026
Safety

Advancing Relevance Measurement with Vision-Language Models for Web-Scale Search

DGX agent

arXiv:2608.02446v1 Announce Type: cross Abstract: Relevance evaluation plays a crucial role in personalized search systems, serving as a guardrail alongside user engagement metrics to ensure that sear

safetyarxiv-cs-lg
4 Aug 2026
Model Releases

Analyzing Speech Condition Effects in Dysarthric ASR: A Layer-wise Probing Study

DGX agent

arXiv:2608.01865v1 Announce Type: new Abstract: Automatic speech recognition (ASR) performance degrades sharply on dysarthric speech, yet how disordered articulation reshapes a model's internal repres

model-releasesarxiv-cs-cl
4 Aug 2026
Local Ai

Beyond Global Latents: Chunk-Based Sparse Grid VAE for Scalable 3D Modeling

DGX agent

arXiv:2608.02016v1 Announce Type: new Abstract: Sparse voxel grids preserve the spatial structure needed for detailed 3D reconstruction, but their memory still grows rapidly with resolution as active

local-aiarxiv-cs-cv
4 Aug 2026
Model Releases

CADENA: Stepwise CAD Reverse Engineering

DGX agent

arXiv:2608.00799v1 Announce Type: new Abstract: Computer-Aided Design (CAD) underpins modern engineering, yet converting existing shapes into editable models still demands substantial expert effort. M

model-releasesarxiv-cs-cv
4 Aug 2026
Research

Can Humans Dream of Electric Sheep? Human-Written Samples for Fine-Grained Vision-and-Language Hallucination Benchmarking

DGX agent

arXiv:2608.01021v1 Announce Type: cross Abstract: In an age of rapid model turnover, how do we make hallucination evaluation more perennial? We explore whether human-written hallucination samples coul

researcharxiv-cs-cl
4 Aug 2026
Research

Can Language Models Identify Shadow Trading Targets? An NLP Evaluation of SEC Enforcement Theory

DGX agent

arXiv:2608.01322v1 Announce Type: new Abstract: Shadow trading -- trading in a peer firm's securities on the basis of material nonpublic information (MNPI) about an 'economically linked' company -- is

researcharxiv-cs-cl
4 Aug 2026
Model Releases

Don't Judge a Book by its Cover: Testing LLMs' Robustness Under Logical Obfuscation

DGX agent

arXiv:2602.01132v2 Announce Type: replace Abstract: Tasks such as solving arithmetic equations, evaluating truth tables, and completing syllogisms are handled well by large language models (LLMs) in t

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

EgoIntent: A Pre-Outcome Micro-Step Benchmark for Understanding What, Why, and Next

DGX agent

arXiv:2603.12147v2 Announce Type: replace Abstract: Egocentric video provides a natural modality for studying human behavior, but conventional visual understanding captures mainly observable scenes, o

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

FinHardBench: Can LLMs Generate Latency-Aware Hardware for Financial Computing?

DGX agent

arXiv:2608.00909v1 Announce Type: new Abstract: Can large language models generate not just correct, but fast hardware? This paper investigates the question in financial FPGA design, where 5-10 nanose

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

From Direction to Magnitude: How Multimodal Instruction-Tuning Reorganizes the Geometric Encoding of Identity-Specifying Prompts in Transformer Hidden States

DGX agent

arXiv:2607.09842v2 Announce Type: replace-cross Abstract: We investigate whether identity-specifying system prompts produce statistically distinguishable geometric fingerprints in the hidden-state tra

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

GIFT: Geometry-Invariant Fine-Tuning for Non-Lambertian Monocular Depth Estimation

DGX agent

arXiv:2608.02068v1 Announce Type: new Abstract: Monocular depth foundation models, benefiting from large-scale synthetic training data, have demonstrated strong generalization. However, they often hal

model-releasesarxiv-cs-cv
4 Aug 2026
Safety

HapticVLA: Contact-Rich Manipulation via Vision-Language-Action Model without Inference-Time Tactile Sensing

DGX agent

arXiv:2603.15257v2 Announce Type: replace Abstract: Tactile sensing is a crucial capability for Vision-Language-Action (VLA) architectures, as it enables dexterous and safe manipulation in contact-ric

safetyarxiv-cs-ro
4 Aug 2026
← Previous
1…275276277278279…1058
Next →