AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
83,773 results
11 Aug 2026

When Does Trace-Driven Evaluation Mislead MoE Expert Caching? Replay Semantics, Workload Contamination, and Operating Regimes

SafetyDGX agent

arXiv:2608.07911v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) models have outgrown accelerator memory, and offloading expert weights to host memory is now standard. This makes expert cache

When Grammar Guides the Attack: Uncovering Control-Plane Vulnerabilities in LLMs with Structured Output

Model ReleasesDGX agent

arXiv:2503.24191v4 Announce Type: replace-cross Abstract: Content Warning: This paper may contain unsafe or harmful content generated by LLMs that may be offensive to readers. Large Language Models (L

When Is a Steerable Concept Representation Real? Measurement Confounds in a Cross-Family Audit of Neuroscience Parallels in LLMs

Local AiDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub

arXiv:2608.08159v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly reported to exhibit human-like neural and cognitive signatures, including concept cells, mental number lin

When Is Benchmark Contamination Detectable? Information Limits and Power-Calibrated Audits

Model ReleasesDGX agent

arXiv:2608.07914v1 Announce Type: new Abstract: Behavioral contamination detectors can return 'no evidence' either because a benchmark is clean or because the audit has little power. We formalize this

When Latents Forget Pixels: Restoring Fidelity in Diffusion Transformer Super-Resolution

TutorialsDGX agent

arXiv:2608.09133v1 Announce Type: cross Abstract: Image super-resolution (SR) with large generative models has recently achieved remarkable perceptual quality, yet maintaining fidelity to the LR obser

When LLM Agents Negotiate: Private Information and Dynamic Bargaining in Supply Chains

Model ReleasesDGX agent

arXiv:2608.07538v1 Announce Type: new Abstract: As LLM agents move from decision support to autonomous procurement, firms need to know whether delegated negotiators create value, divide it predictably

When should we trust the annotation? Selective prediction for molecular structure retrieval from mass spectra

Model ReleasesDGX agent

arXiv:2603.10950v2 Announce Type: replace Abstract: Machine learning methods for identifying molecular structures from tandem mass spectra (MS/MS) have advanced rapidly, yet current approaches still e

When Skills Meet Safety: Benchmarking and Characterizing the Adaptive Jailbreak Robustness of Skill-Merged LLMs

Model ReleasesDGX agent

arXiv:2608.08542v1 Announce Type: new Abstract: Model merging has become the default way to give an aligned language model new skills without retraining: a practitioner folds task vectors from math, c

When the Judge Should Not Decide: Evidence-Locked, Non-Compensatory Selection Bounds LLM-Judge Failure in Reasoning Pipelines

Model ReleasesDGX agent

arXiv:2608.07813v1 Announce Type: new Abstract: An LLM judge deployed inside a reasoning pipeline does not merely measure quality, it decides which answer ships. We show that the cost of that decision

Where Is the Bee? Detecting Tiny Pollinators with a Single Collaborative-Head Transformer

ResearchDGX agent

arXiv:2608.08580v1 Announce Type: new Abstract: The CVPPA@ECCV 2026 BuzzSpot Challenge asks us to detect bees, bumblebees, hoverflies, and moths in 1920x1080 field keyframes. Its annotations carry 2 d

Who Bridges Safety? Identifying and Targeting Cross-Lingual Shared Safety Pathways

Local AiDGX agent

arXiv:2608.09095v1 Announce Type: new Abstract: Uncovering the internal mechanisms underlying the safety capabilities of large language models (LLMs) is crucial for developing trustworthy artificial i

Who Built This Model? Tracing LLM Lineage via Spectral Fingerprints in Weight Space

SafetyDGX agent

arXiv:2608.07786v1 Announce Type: new Abstract: Open-weight large language models (LLMs) are increasingly developed through complex, multi-stage pipelines, leading to intricate lineage relationships t

Who Verifies the Benchmark? Decentralizing Trust in Large Language Model Evaluation

Model ReleasesDGX agent

arXiv:2608.07762v1 Announce Type: new Abstract: LLM benchmarks can build an organization's reputation and attract customers, but only when results are transparent and verifiable. Unverified claims tha

Whoa, Meta released a new open-weight LLM yesterday, something that hasn't happened since the good old Llama days. Their Meta Muse Glimmer m…

Model ReleasesDGX agent

Whoa, Meta released a new open-weight LLM yesterday, something that hasn't happened since the good old Llama days. Their Meta Muse Glimmer model is a 30B multimodal reasoning model with a Gemma-like a

Why Does the Future Branch? Identifiable Closure Tests for Stochastic Physical World Models

Model ReleasesDGX agent

arXiv:2608.00591v2 Announce Type: replace Abstract: A calibrated stochastic world model can reveal how uncertain a future is without revealing why it branches. The same conditional future law can aris

Why Scaling AI Compute Performance Requires a New Power Architecture

HardwareDGX agent

Every new generation of accelerated computing demands more from the infrastructure underneath it — more compute performance, higher rack density and more efficient, scalable power distribution. The bo

Wiener Representation Filtering for VLM Hallucination Suppression

Model ReleasesDGX agent

arXiv:2608.08167v1 Announce Type: new Abstract: Vision-language models (VLMs) excel at open-ended captioning and visual QA but often describe objects, attributes, or relations absent from the image, a

Wisdom in Unity: The Role of Multilingual Training in Figurative Language Identification in Proverbs

ResearchDGX agent

arXiv:2608.08090v1 Announce Type: new Abstract: Although multilingual approaches to figurative language identification are not new, the shift beyond language homogeneous training data requires a clear

Wix launches Symphony, a new standalone multi-agent system built for business operations

Model ReleasesDGX agent

Cloud-based website builder Wix Ltd. today announced the launch of Symphony, a new standalone agentic artificial intelligence platform that proactively learns business values, interests, needs, practi

World Simulator: Queer Erotica and the Absurdity of AI Video Models That Promise the World

ResearchDGX agent

arXiv:2608.07510v1 Announce Type: cross Abstract: Increasingly, AI video models are marketed as 'world simulators,' suggesting their ability to model infinite realities. Despite such claims, these mod

World Tokens: Enhancing Embodied Policies with Training-Time World Modeling

SafetyDGX agent

arXiv:2608.09730v1 Announce Type: new Abstract: Vision-language-action (VLA) models are a widely adopted paradigm for embodied policies. They excel at efficient closed-loop control but do not explicit

WorldSimProbe: Diagnosing Simulator Faithfulness in Action-Conditioned World Models for Embodied Manipulation

Model ReleasesDGX agent

arXiv:2608.09298v1 Announce Type: cross Abstract: Action-conditioned world models (ACWMs) promise to provide embodied AI with scalable predictive simulators for planning, policy evaluation, and data g

WRAP: Wasserstein-Robust Adaptive Plug-in for Robot Localization

Local AiDGX agent

arXiv:2608.09807v1 Announce Type: new Abstract: Robotic localization under changing sensing conditions can suffer from biased errors and miscalibrated covariances. We present WRAP, an adapter-agnostic

Writing formats I can no longer read because AI has beaten them to death

TutorialsDGX agent

I don’t even care anymore whether these posts are *actually* AI-generated. The problem is that there’s now a very specific style of internet writing that instantly makes my brain refuse to continue re

WuYuEval: A Multi-Level Benchmark for Large Language Models in Solid Waste Management

Model ReleasesDGX agent

arXiv:2608.07529v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as technical assistants, but their competence in solid waste management (SWM) remains difficult to

X2C: A Dataset Featuring Nuanced Facial Expressions for Realistic Humanoid Imitation

Model ReleasesDGX agent

arXiv:2505.11146v3 Announce Type: replace-cross Abstract: Fine-grained facial expression transfer from humans to humanoid agents presents a unique pattern recognition challenge due to the significant

xAI co-founder Igor Babuschkin's River AI raised $1B led by General Catalyst to build home or small business computer servers capable of running AI locally (Cade Metz/New York Times)

Local AiDGX agent

Cade Metz / New York Times: xAI co-founder Igor Babuschkin's River AI raised $1B led by General Catalyst to build home or small business computer servers capable of running AI locally — Igor Babuschki

XClipGS: Exact Half-Space Clipping for Medical Volume Gaussian Splatting

Local AiDGX agent

arXiv:2608.07760v1 Announce Type: new Abstract: Gaussian-splatting proxies enable interactive rendering of volumetric medical scans, but a clipping plane exposes anatomy not constrained by external-vi

XEns-CKD: An Explainable Ensemble-Based Approach for Chronic Kidney Disease Stage Detection

ResearchDGX agent

arXiv:2608.07561v1 Announce Type: new Abstract: Chronic kidney disease (CKD) is a silent disease. Its progression may not significantly hamper a person's daily routine. Human kidney function can be cl

XFeat Revisited: Reproducibility and Evaluation of a Lightweight Image Matcher

Model ReleasesDGX agent

arXiv:2608.09519v1 Announce Type: new Abstract: We present a reproducibility study of XFeat, a lightweight local feature extractor and matcher designed to identify corresponding points across images e

XPolicyLab: A Unified Standard and Open Ecosystem for Robot Policy Evaluation and Deployment

Model ReleasesDGX agent

arXiv:2608.09892v1 Announce Type: new Abstract: Robot policy evaluation and deployment remain fragmented by model-specific software dependencies, data representations, and runtime interfaces, so that

Yesterday's Shield, Today's Spear: A Self-Evolving Safety Guardrail in Production

SafetyDGX agent

arXiv:2608.08471v1 Announce Type: new Abstract: Deployed LLM safety guardrails are predominantly static: trained once and frozen at release, while new jailbreak techniques and previously un-addressed

You Don't Need To Stay in The Loop: An Agentic Robotics Loop for Robot-Policy Improvement

Model ReleasesDGX agent

arXiv:2608.07555v1 Announce Type: new Abstract: Coding agents such as Claude Code and Codex close the software loop: a main agent manages the loop, subagents analyze and execute, tools do the work. We

You Only Flow Once: Calibrated and Real-Time Radar Pose Estimation with Multi-Hypothesis Normalizing Flows

ResearchDGX agent

arXiv:2608.09579v1 Announce Type: new Abstract: Sparse and noisy millimeter-wave radar point cloud observations often correspond to multiple plausible human poses, making deterministic pose estimation

Your Prompt Is Not the Only Prompt: How Much Do LLMs Weight Structured-Output Schema Descriptions?

Model ReleasesDGX agent

arXiv:2608.08254v1 Announce Type: new Abstract: Structured output, where an LLM populates a predefined JSON schema, has become a default mechanism for data labeling and information extraction, but it

Your VLM Already Knows When: Training-Free Temporal Grounding by Asking Yes or No

Model ReleasesDGX agent

arXiv:2608.08315v1 Announce Type: new Abstract: Multimodal LLMs that recognise events reliably still fail to say when they happen. Prompted for timestamps, strong VLMs reach as little as 3.8% R@0.5 on

Zero-shot 2D Grounding with Novel Affordance Types

Model ReleasesDGX agent

arXiv:2608.08929v1 Announce Type: new Abstract: 2D affordance grounding aims to locate the region of an object that a human can interact with. Existing research focuses on recognizing affordance types

Zero-Shot Traffic Accident Detection via a Coarse-to-Fine VLM-Tracking Pipeline

Model ReleasesDGX agent

arXiv:2608.08867v1 Announce Type: new Abstract: Traffic surveillance cameras capture accidents continuously, yet converting raw CCTV footage into structured event records that pinpoint when, where, an

ZeroLock: Concurrent Memory-Efficient LLM Training via Modular Update Decoupling

Local AiDGX agent

arXiv:2608.07974v1 Announce Type: new Abstract: Large language model (LLM) fine-tuning at the edge adapts the model to scenario-specific data while preserving privacy. Although existing studies propos

ZetaGPT: A Reference Implementation of Positional--Encoding--Free State--Space--Attention Language Models

ResearchDGX agent

arXiv:2608.09432v1 Announce Type: cross Abstract: Transformer-based language models rely on self-attention, whose computation is permutation-equivariant and therefore lacks an intrinsic mechanism for

ZhuLong: Execution-Grounded LLM Agent for EDA Scripting with Offline API Self-Exploration

Model ReleasesDGX agent

arXiv:2608.07925v1 Announce Type: new Abstract: EDA scripting with tool-specific, often undocumented APIs remains a long-tail bottleneck that existing LLMs fail to address. This paper presents ZhuLong

ZOMP: Zeroth-Order Multi-Modal Prompt Tuning for Vision-Language Models

ResearchDGX agent

arXiv:2608.08060v1 Announce Type: new Abstract: Fine-tuning vision-language models such as CLIP typically requires backpropagation (BP) through the full model, which is infeasible when only forward-pa

10 Aug 2026

1M context with 17 GB model in 24 GB VRAM: 'for the first time I was able to load a context of almost 1M tokens and extract 7 needles from various parts of the text'

Model ReleasesDGX agent

https://preview.redd.it/xxjh11f38jih1.png?width=1852&format=png&auto=webp&s=76850ed51e29a8bc86c2ca718d4320075eed4363 Just wanted to share a user report that I found to be very interesting. Some person

2/ New deep dive: Autoscaling endpoints for LLM inference. Dedicated Inference can scale on eight metrics. inflight_requests is the default …

HardwareDGX agent

2/ New deep dive: Autoscaling endpoints for LLM inference. Dedicated Inference can scale on eight metrics. inflight_requests is the default because it sees queue pressure before latency degrades. We t

8 days ago, while jogging, I asked Claude to solve the Riemann Hypothesis It didn’t. 1.5 days later, it proved >= 67% of the zeros are on th…

Model ReleasesDGX agent

8 days ago, while jogging, I asked Claude to solve the Riemann Hypothesis It didn’t. 1.5 days later, it proved >= 67% of the zeros are on the line (prev: 41.6%) Still not sure what that means, but som

A Disturbance in the Force: Force Actuation on the RAVEN II Surgical Robot with Parallel Motor-Cable Units

ResearchDGX agent

arXiv:2608.06488v1 Announce Type: new Abstract: Difficulty in haptic feedback for surgical robots has been a long-term problem for decades. In recent years, learning-based force estimation from robot

A downside with VLM-based parsing is that they’re generally slower than text-based heuristic approaches. As a result they add latency to any…

Model ReleasesDGX agent

A downside with VLM-based parsing is that they’re generally slower than text-based heuristic approaches. As a result they add latency to any ad-hoc file processing *in-the agent loop* (e.g. if you upl

A Finite E-Group of Nilpotency Class Three

ResearchDGX agent

arXiv:2608.07275v1 Announce Type: cross Abstract: A group is an E-group if every element commutes with each of its endomorphic images. Caranti asked whether a finite E-group can have nilpotency class

A foundation-model approach to pediatric headache classification from rs-fMRI

ResearchDGX agent

arXiv:2608.07287v1 Announce Type: new Abstract: Headache is the most common neurological disorder in children and substantially affects quality of life. We investigated whether resting-state functiona

A Haptic Robot Finger Designed for Guqin Instrument Playing

ResearchDGX agent

arXiv:2608.07002v1 Announce Type: new Abstract: With the rapid advancement of humanoid robotics and embodied intelligence technologies, numerous musical instrument-playing robots have emerged in recen

A look at London's King's Cross, which transformed from a seedy area into an AI hub after DeepMind arrived in 2016, and now hosts OpenAI, Meta, Wayve, and more (Dominic-Madori Davis/TechCrunch)

IndustryDGX agent

Dominic-Madori Davis / TechCrunch: A look at London's King's Cross, which transformed from a seedy area into an AI hub after DeepMind arrived in 2016, and now hosts OpenAI, Meta, Wayve, and more — Wha

A MARL Centered Reference Architecture for Large Language Model Augmentation in Smart Manufacturing

Local AiDGX agent

arXiv:2608.07148v1 Announce Type: new Abstract: Modern manufacturing imposes six coupled demands on adaptive control: local decisions with global consequences, partial observability, nonstationarity,

A Multi-Agent Framework for Automated Coarse-Grained Molecular Dynamics of Polymers

Model ReleasesDGX agent

arXiv:2608.06694v1 Announce Type: new Abstract: Coarse-grained (CG) molecular dynamics extends polymer simulation beyond the scales accessible to all-atom (AA) methods, but bottom-up CG modeling is la

A Physics-Inspired Classical Digital Twin of Cortical Dynamics: A Band-Stratified Metriplectic Port-Hamiltonian Neural Network Learned from Brain-Computer-Interface EEG

ResearchDGX agent

arXiv:2607.10439v3 Announce Type: replace-cross Abstract: We present a physics-inspired classical digital twin of brain-computer- interface (BCI) data: a graph neural network constrained to a band-str

A Picture is Worth a Thousand Tokens: How Vision Language Models Cut AI Energy Costs While Improving Accuracy

Model ReleasesDGX agent

arXiv:2608.07427v1 Announce Type: new Abstract: LLM inference accounts for over 90% of AI operational energy, scaling directly with input token count---a critical inefficiency for telecom network anal

A Practical Evaluation Method for Long-Form Simultaneous Speech-to-Speech Translation

SafetyDGX agent

arXiv:2606.15059v2 Announce Type: replace Abstract: Simultaneous speech-to-speech translation (SimulS2ST) enables real-time cross-lingual communication, but existing evaluation has focused largely on

A primer on optimal transport for causal inference with observational data

ResearchDGX agent

arXiv:2503.07811v3 Announce Type: replace-cross Abstract: The theory of optimal transportation has developed into a powerful and elegant framework for comparing probability distributions, with wide-ra

A profile of Founders Pledge, which has recorded $4B in pledges from founders so far in 2026, as the AI industry now accounts for a third of its lifetime total (Joel Khalili/Wired)

IndustryDGX agent

Joel Khalili / Wired: A profile of Founders Pledge, which has recorded $4B in pledges from founders so far in 2026, as the AI industry now accounts for a third of its lifetime total — A new generation

A proximal subgradient method for nonconvex stochastic optimization under the Kurdyka-{L}ojasiewicz condition

ResearchDGX agent

arXiv:2608.05460v1 Announce Type: cross Abstract: This work introduces a proximal stochastic subgradient method for minimizing the sum of an expected cost, whose integrand is potentially nonsmooth and

A Rate Separation for Agnostic Direct Sums

ResearchDGX agent

arXiv:2608.06951v1 Announce Type: new Abstract: Hanneke, Moran, and Waknine ite{HannekeMoranWaknine2024} asked how the agnostic PAC learning curve of the direct sum C^r depends on the single-instance

← Previous
1…4445464748…1397
Next →