AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,457
  • Agents7,399
  • Applications5,302
  • Concepts5
  • Hardware1,786
  • Industry6,117
  • Local Ai4,835
  • Model Releases23,193
  • Research19,715
  • Safety13,094
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,457
  • Agents7,399
  • Applications5,302
  • Concepts5
  • Hardware1,786
  • Industry6,117
  • Local Ai4,835
  • Model Releases23,193
  • Research19,715
  • Safety13,094
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

86,457Total entries
1Added by human
86,456Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,039 results
10 Aug 2026

Multi-Agent Forensic Reasoning for Generalizable Deepfake Video Detection

Model ReleasesDGX agent

arXiv:2608.06865v1 Announce Type: cross Abstract: The malicious use of generative artificial intelligence to create highly realistic deepfake videos raises serious ethical concerns and poses substanti

Pre-Inference Routing for Cost-Efficient Document Field Extraction

ResearchDGX agent

arXiv:2608.06607v1 Announce Type: new Abstract: Most document-extraction systems use a single model for all documents. This is simple but can be costly for easy cases and less effective for difficult

ReQuant: Fixed-Grid Discrete Refinement for Post-Training Quantization

ResearchDGX agent

arXiv:2608.07019v1 Announce Type: new Abstract: Post-training quantization (PTQ) is widely used to reduce the memory and computational cost of large language models. Existing PTQ methods typically obt

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Stoicheia: Character-Level Masked Diffusion for Ancient Greek Textual Restoration, Parsing, and Metrical Scansion

Model ReleasesDGX agent

arXiv:2608.07249v1 Announce Type: new Abstract: We introduce Stoicheia, a 405M-parameter character-level masked-diffusion encoder for Ancient Greek whose input factors into five aligned, independently

Tensor Network Kernel Machines: A JAX Framework for Machine Learning and Nonlinear System Identification

Model ReleasesDGX agent

arXiv:2608.07043v1 Announce Type: cross Abstract: Developing nonlinear models that are both expressive and computationally efficient remains a challenge in machine learning and nonlinear system identi

9 Aug 2026

CyberKimi just dropped strong results on one of ExploitBench’s hardest V8 bugs , points away from Mythos

Model ReleasesDGX agent

Hey everyone ! Quick share from the cyber + local LLM side of things that I found interesting. During this week’s hacker summer camp, an AI researcher and reverse malware engineer veteran 'lordx64' on

KLQ: Training-free measured rotation quantization. Beats all training-free rotation-based quantization methods on W4A4KV4-bits. Llama 3.2 1B KLQ-quantized beats SpinQuant and gets close to ReSpinQuant without GPTQ/LDLQ rounding.

Model ReleasesDGX agent

First of all, I'm not a lab, this was a solo summer research project that finally culminated into the github repo and the writeup. The repo includes a much deeper dive with methods, findings about qua

7 Aug 2026

~45% lower MiniMax H3 sampler time with new Spectrum settings — degree 1 works surprisingly well (v0.1.8)

Model ReleasesDGX agent

Follow-up to my original Spectrum MiniMax H3 post: https://www.reddit.com/r/StableDiffusion/comments/1vf1ze3/spectrum_acceleration_for_minimax_h3_in_comfyui/ In that first post, I released the MiniMax

Am I just hallucinating

Model ReleasesDGX agent

Or is there any reason why I feel like model output quality seems to be better when I use higher micro-batch values (ub) in llama-cpp? I don't really have any hard numbers or anything (just running th

Arbitrage: Efficient Reasoning via Advantage-Aware Speculation

ResearchDGX agent

Modern Large Language Models achieve impressive reasoning capabilities with long Chain of Thoughts, but they incur substantial computational cost during inference, and this motivates techniques to imp

Basically every remaining good AI benchmark score has an implied asterisk next to it which reads: * could be signficantly higher with a bett…

Model ReleasesDGX agent

On August 7, 2026 Ethan Mollick tweeted that “every remaining good AI benchmark score has an implied asterisk next to it which reads: * could be significantly higher with a better harness.” The commen

BendTwin: Robust Dense-to-Sparse Physical Reconstruction with Bending-Aware Differentiable Spring-Mass Models

ResearchDGX agent

arXiv:2608.06164v1 Announce Type: new Abstract: Reconstructing objects with mechanical properties from video observations enables physically consistent dynamic prediction, benefiting robotics planning

Decoupling Perception from Description: Computation-Grounded Representation Alignment between Multivariate Time Series and Language

Model ReleasesDGX agent

arXiv:2608.05238v1 Announce Type: new Abstract: Training multimodal models to align time series with language runs into a self-supervision trap. The usual recipe asks an LLM to read a series and write

Evaluating and Improving Pedagogical Fit in LLM-Based AI Tutors with the Pedagogical Suitability Index

Model ReleasesDGX agent

arXiv:2608.05411v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used as AI tutors, but a correct answer is not always a pedagogically appropriate one. In classroom learni

FormBharo: Designing and Evaluating a Voice Agent for Conversational Form Filling in Rural India

Model ReleasesDGX agent

arXiv:2608.06027v1 Announce Type: cross Abstract: In India, almost every social benefit starts with a form, yet the people who need these benefits most are often unable to read or write. Reaching them

Hijacking Robots with a Piece of Paper: A Systematic Study of Physical Prompt Injection in VLM-Controlled Robots

Model ReleasesDGX agent

arXiv:2608.05715v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) are increasingly deployed as planners in robotic systems, where they translate natural-language commands into executable

Kastor: An efficient fine-tuning strategy for generative emulation of PDE simulations

Model ReleasesDGX agent

arXiv:2608.06107v1 Announce Type: new Abstract: Machine learning offers a promising avenue to accelerate physical simulations by replacing computationally expensive traditional Partial Differential Eq

LangChoiceBench: Measuring and Explaining Programming-Language Choice in LLMs

Model ReleasesDGX agent

arXiv:2608.06041v1 Announce Type: cross Abstract: Large language models (LLMs) have been shown to exhibit strong Python preferences when generating project-level code, but there is currently no system

Look Twice: Training-Free Evidence Highlighting for Knowledge-based Visual Question Answering

Model ReleasesDGX agent

arXiv:2604.01280v2 Announce Type: replace-cross Abstract: Knowledge-based Visual Question Answering (KB-VQA) requires Multimodal Large Language Models (MLLMs) to identify and combine fine-grained visu

Measuring and Detecting Harmful AI Sycophancy

TutorialsDGX agent

arXiv:2608.05624v1 Announce Type: new Abstract: Sycophantic responses are becoming pervasive in large language models (LLMs), and prior work has pointed out that some of them could be harmful. This pa

On-Policy Self-Distillation without Any Supervision

SafetyDGX agent

arXiv:2608.06296v1 Announce Type: new Abstract: On-policy (Self-)Distillation (OPD / OPSD) has shown strong potential for post-training large language models (LLMs). However, existing methods still re

OPERA: Operator-residual feedback for reliable autonomous optical experiments with language-model agents

AgentsDGX agent

arXiv:2608.05990v1 Announce Type: new Abstract: Autonomous agents choose actions using scores that may not reflect experimental success. We developed OPERA, an operator-residual framework for optical

Scaling Categorical Flow Maps

ResearchDGX agent

Continuous diffusion and flow matching models could represent a powerful alternative to autoregressive approaches for language modelling (LM), as they unlock a host of advantages currently reserved fo

Seeing Is Not Deciding: Can Multimodal LLMs Act as Effective CEOs?

Model ReleasesDGX agent

arXiv:2608.05864v1 Announce Type: new Abstract: Large language models are increasingly applied as autonomous decision-making agents. However, in executive business decisions, existing benchmarks are l

Serving Deepseek v4 Flash 0731 on 2x DGX Spark — 5-7 GB OS headroom, what would you do to lower VRAM usage and increase OS available RAM?

Model ReleasesDGX agent

Hey all, I'm serving DSv4Flash 0731 on a cluster of 2x DGX Sparks but am running into constant issues with having almost no RAM (unified memory) left for the OS/cache and I'd love to hear the communit

Shrinking the Generation-Verification Gap with Weak Verifiers

Model ReleasesDGX agent

arXiv:2506.18203v3 Announce Type: replace Abstract: Verifiers can improve language model capabilities by scoring and ranking responses from generated candidates. Currently, high-quality verifiers are

Spectral Distillation: From Nonlinear Dynamics to Linear State-Space Models

TutorialsDGX agent

arXiv:2608.05416v1 Announce Type: new Abstract: Can nonlinear dynamical systems be learned through a compact linear state-space representation, without directly solving a non-convex system-identificat

Subliminal Learning is Non-Semantic Distillation

Model ReleasesDGX agent

arXiv:2608.05734v1 Announce Type: new Abstract: Subliminal Learning (SL) is a surprising type of generalization displayed by modern language models. It allows the transfer of a bias or behavior from a

Thanks for the thorough testing! With Qwen3.8-Max, everyone can observe the world in detail. 👀

Model ReleasesDGX agent

Alibaba’s Qwen team announced the new Qwen‑3.8‑Max model and thanked users for extensive testing, noting its ability to provide highly detailed world observations. A community member highlighted that

Why the Third Axis Is Freedom

ResearchDGX agent

arXiv:2608.05423v1 Announce Type: cross Abstract: In generative training, a model produces an output and is penalised for its difference from an example. With one output per comparison, a model that p

6 Aug 2026

Attention, Anomalies! Handling Attention Layers in Unsupervised Federated Outlier Detection

ResearchDGX agent

arXiv:2608.04753v1 Announce Type: new Abstract: Attention layers are the backbone of today's most powerful and impactful models. Models with multi-million and billion parameters rely on contextual kno

Can Post-Training Transform LLMs into Causal Reasoners?

Model ReleasesDGX agent

arXiv:2602.06337v2 Announce Type: replace-cross Abstract: Causal inference is essential for decision-making but remains challenging for non-experts. While large language models (LLMs) show promise in

CoCo-IR: Contextual Composed Image Retrieval

Model ReleasesDGX agent

arXiv:2608.05149v1 Announce Type: new Abstract: Current instruction-based image retrieval systems are powerful but limited to single-turn interactions, failing to capture the iterative nature of compl

ContextMaster: Interactive Multi-Shot Video Creation via Fixed-Budget Sparse Context Routing

HardwareDGX agent

arXiv:2608.04956v1 Announce Type: new Abstract: Recent video models increasingly support generation, reference conditioning, and editing within a single model, yet typically expose them as separate op

Echo Flow Networks

Model ReleasesDGX agent

arXiv:2509.24122v3 Announce Type: replace Abstract: At the heart of time-series forecasting (TSF) lies a fundamental challenge: how can models efficiently and effectively capture long-range temporal d

FinPerMA: A Theory-Informed, Event-Grounded Personalized-Memory Benchmark for LLM Agents

Model ReleasesDGX agent

arXiv:2608.04095v1 Announce Type: new Abstract: Large language model (LLM) agents are increasingly used as personalized assistants in high-stakes domains such as financial advising, yet it remains unc

FinReportBench: Measuring and Improving Institution-Grade Financial Report Generation

Model ReleasesDGX agent

arXiv:2608.04374v1 Announce Type: cross Abstract: Large language models can produce fluent financial analysis, but fluency alone does not establish whether a report is suitable for institutional deliv

K-EXAONE 2.0 Technical Report

Model ReleasesDGX agent

arXiv:2608.04505v1 Announce Type: new Abstract: This technical report presents K-EXAONE 2.0, an open-weight multilingual foundation model developed by LG AI Research as a step in our effort toward glo

Kathleen Writes: Autoregressive Generation and Data Scaling Without Attention

Model ReleasesDGX agent

arXiv:2608.04678v1 Announce Type: new Abstract: Papers 1-2 of the Kathleen series showed that a byte-level, attention-free architecture built from a wavetable encoder and multi-scale reverberant state

Locking Pretrained Weights via Deep Low-Rank Residual Distillation

ResearchDGX agent

The quality of open-weight language models has dramatically improved in recent years. Sharing weights greatly facilitates model adoption by enabling their use across diverse hardware and software plat

OmniEdit-Bench: A Comprehensive Benchmark for Instruction-based Video Editing

Model ReleasesDGX agent

arXiv:2608.05049v1 Announce Type: new Abstract: Instruction-based video editing (IVE) is an emerging field with broad applications, yet evaluating editing models remains challenging. Existing benchmar

Reading Between the Frames: Interpreting Implicit and Non-literal Meaning in Social Media Videos

Model ReleasesDGX agent

arXiv:2608.04939v1 Announce Type: new Abstract: Social media videos often communicate meanings that go beyond their visible actions, captions, or speech. A mundane clip may become humorous, ironic, or

RepoProbe: Benchmarking Architecture-Aware Repository Comprehension with Checklists

Model ReleasesDGX agent

arXiv:2608.04783v1 Announce Type: cross Abstract: The integration of Large Language Models (LLMs) into software engineering has shifted the focus from function-level generation to repository-scale ass

RooflineBench: A Benchmarking Framework for On-Device LLMs via Roofline Analysis

Model ReleasesDGX agent

arXiv:2602.11506v4 Announce Type: replace-cross Abstract: The transition toward localized intelligence through Small Language Models (SLMs) has intensified the need for rigorous performance characteri

SEAR: Simple and Efficient Adaptation of Visual Geometric Transformers for Unpaired RGB+Thermal 3D Reconstruction

Model ReleasesDGX agent

arXiv:2603.18774v2 Announce Type: replace Abstract: Foundational feed-forward visual geometry models enable accurate and efficient camera pose estimation and scene reconstruction by learning strong sc

Social Pressure Breaks Majority Voting in LLM Safety Panels

SafetyDGX agent

arXiv:2608.04415v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used to detect unsafe content. A common approach is to combine judgments from a panel of models to correct

Teaching Nemotron Greek: Mining a Corpus, Adapting Retrieval, and Grounding Generation for Modern Greek across Specialist Domains

Model ReleasesDGX agent

arXiv:2608.05138v1 Announce Type: cross Abstract: Modern Greek is absent from NVIDIA's Nemotron retrieval models and from major multilingual retrieval benchmarks, despite being important for retrieval

Text2GraphQuery-Bench: A Text to Graph Query Benchmark

Model ReleasesDGX agent

arXiv:2602.11745v2 Announce Type: replace Abstract: Graph models are fundamental to data analysis in domains rich with complex relationships. Unlike SQL, which benefits from a rel- atively unified sta

When Prompts Become Pixels: Prompt-Region Grounding for Multimodal Reasoning

Model ReleasesDGX agent

arXiv:2608.04726v1 Announce Type: new Abstract: Multimodal large language models increasingly reason over screenshots and documents where the task itself may be written in pixels. Yet benchmarks usual

5 Aug 2026

AI Security Leaderboard: Methodology, Results and Minimal Standard

Model ReleasesDGX agent

arXiv:2608.03070v1 Announce Type: cross Abstract: Frontier AI model developers increasingly rely on layered safeguards to prevent catastrophic misuse, but little public evidence exists on how much pro

ChartAnno: Evaluating MLLMs for Chart Annotation Generation

Model ReleasesDGX agent

arXiv:2608.03464v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have made significant progress in chart understanding, generation, and editing, but their ability to annotate e

DenialRAG: Single-Document RAG Poisoning via Embedded Parametric Denial

Model ReleasesDGX agent

arXiv:2608.02678v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) systems are vulnerable to corpus poisoning: an attacker who inserts a crafted document into the retrieval corpus

EduClaw-Bench: A Long-Horizon Benchmark for Pedagogical LLM Agents with Simulated Learners

Model ReleasesDGX agent

arXiv:2608.03206v1 Announce Type: cross Abstract: Large language models (LLMs) power educational applications from tutoring to essay scoring, but each is a point solution to a single task, and only re

FinVerse: Financial Time-Series Benchmark

Model ReleasesDGX agent

arXiv:2608.03259v1 Announce Type: cross Abstract: As time-series foundation models have emerged, the need for benchmarks that can evaluate their forecasting ability in meaningful ways has become incre

I took a local OCR model's accuracy from 60% to 99%

Local AiDGX agent

I built a local OCR pipeline a few days ago, and it turned into a surprisingly interesting experiment—taking accuracy from around 60% to 99%. I wrote a short blog about what worked, what failed, and t

Improved Quantum Algorithms for Reinforcement Learning Under a Generative Model

SafetyDGX agent

arXiv:2608.02826v1 Announce Type: cross Abstract: Reinforcement learning is a subfield of machine learning that studies how an agent interacts with an environment in order to extract as large a reward

LACE: Large Language Model Aided Multi-Agent Framework for Agile RISC-V Instruction Extension

AgentsDGX agent

arXiv:2608.02915v1 Announce Type: cross Abstract: Domain-specific Instruction Set Architecture eXtensions (ISAX) are widely adopted in the RISC-V ecosystem to accelerate emerging workloads, but implem

LFM2.5-2.6B on a OnePlus 13 at 17 tok/s ~ Pure CPU

Model ReleasesDGX agent

As you all know the model is 2.69B parameters with a 128K context window and purpose-built for multi-step agent workflows. What you are seeing is the Q4_K_M GGUF running on my own inference engine bui

MaterialFusion: High-Quality, Zero-Shot, and Controllable Material Transfer with Diffusion Models

ApplicationsDGX agent

arXiv:2502.06606v3 Announce Type: replace Abstract: Manipulating the material appearance of objects in images is critical for applications like augmented reality, virtual prototyping, and digital cont

PASE: Leveraging the Phonological Prior of WavLM for Low-Hallucination Generative Speech Enhancement

ResearchDGX agent

arXiv:2511.13300v1 Announce Type: cross Abstract: Generative models have shown remarkable performance in speech enhancement (SE), achieving superior perceptual quality over traditional discriminative

← Previous
1…270271272273274…1034
Next →