AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
84,460 results
5 Aug 2026

Cross-Anesthetic ECoG State Decoding Fails at the Decision Threshold, Not the Representation

ResearchDGX agent

arXiv:2608.02646v1 Announce Type: cross Abstract: Decoders of anesthetic state from cortical activity fail across drug classes, most notoriously ketamine, but reported accuracy cannot say whether the

CROSS: Cascaded Distillation and Dual-Constraint Grounding for Remote Sensing Referring Segmentation

SafetyDGX agent

arXiv:2608.03147v1 Announce Type: new Abstract: Referring Remote Sensing Image Segmentation (RRSIS) has achieved significant progress through the integration of VLMs and the Segment Anything Model (SA

Cross-Country Learning for National Infectious Disease Forecasting Using European Data

ApplicationsDGX agent

arXiv:2601.20771v2 Announce Type: replace-cross Abstract: Accurate forecasting of infectious disease incidence is critical for public health planning and timely intervention. While most data-driven fo

Content type
AllBlogX PostPaperYouTubeRedditGitHub

Cross-Layer Interaction under Weight-Space Ablation: A Closed-Form Attention Jacobian Bound and a Test on a Real Pretrained Model

ResearchDGX agent

arXiv:2608.03629v1 Announce Type: new Abstract: A companion paper studies when activation patching and weight-space ablation agree, inside an idealized model where a conditional computation is carried

Cross-Lingual Bias in Large Language Models: A Comparative Analysis of English and Swahili

Model ReleasesDGX agent

arXiv:2608.03532v1 Announce Type: new Abstract: Large language models are increasingly deployed in multilingual contexts, yet safety alignment and bias evaluation remain overwhelmingly English-centric

Cross-Model KV Cache Transfer in LLM Families: A Closed-Form Linear Mapping for Prefill Reuse

ApplicationsDGX agent

arXiv:2608.03893v1 Announce Type: new Abstract: Production deployments often swap between different-sized models in a family for cost-quality cascading, mid-conversation switching, and routing, and ea

CrossScope: A Role-Asymmetric World Model for Joint Dual-Scope Surgical Video Prediction

Model ReleasesDGX agent

arXiv:2608.03211v1 Announce Type: new Abstract: Visual world models typically learn future dynamics from a single observation stream, limiting their ability to model cooperative systems with multiple

CRS-Triage: Confidence- and Reliability-Aware Selective Triage under Incomplete Clinical Evidence

ResearchDGX agent

arXiv:2608.03862v1 Announce Type: new Abstract: Emergency triage requires reliable decisions within a short time period. However, the available electronic health record (EHR) data, including structure

CT-HEG: A Bidirectional, Timestamp-Attributed Event Graph for ICU In-Hospital Mortality Prediction - An Architectural Ablation Study

SafetyDGX agent

arXiv:2608.02663v1 Announce Type: cross Abstract: Accurate ICU mortality prediction requires modeling irregular clinical observations across heterogeneous entity types. Existing sequence models handle

CUADebug: Diagnosing and Repairing Computer-Use Agent Failures

Model ReleasesDGX agent

arXiv:2608.02643v1 Announce Type: cross Abstract: Computer-use agents (CUAs) operate real desktop and web interfaces through screenshots, mouse and keyboard actions, and stateful UI feedback, yet thei

CUDA MPC: A GPU-Native Solver for Model Predictive Control

Model ReleasesDGX agent

arXiv:2608.03051v1 Announce Type: new Abstract: Model Predictive Control (MPC) delivers constraint-aware control, but its reliance on online optimization limits its use on systems with fast dynamics,

Cura 1T: Specialized Model for Agentic Healthcare

Model ReleasesDGX agent

arXiv:2607.15314v2 Announce Type: replace Abstract: Healthcare AI agents handle patient consultation, clinical reasoning over text and images, interactive diagnosis, and electronic health record (EHR)

CURV: Enhancing Chart Understanding Through Curriculum Visual Grounded Reasoning

ApplicationsDGX agent

arXiv:2608.02833v1 Announce Type: cross Abstract: Chart question answering (CQA) requires multimodal large language models (MLLMs) to integrate visual comprehension with logical reasoning, yet current

CVPO: Enhancing LLM Reinforcement Learning Reasoning via Value-Variance Adaptation and Dynamic Curriculum Learning

SafetyDGX agent

arXiv:2608.03068v1 Announce Type: cross Abstract: Reinforcement learning (RL) has emerged as an effective method for enhancing the reasoning capabilities of large language models (LLMs). However, exis

DAIF: A Data-Driven Intermediate Fusion Framework for Multimodal Supervised Learning via Approximate Message Passing

Model ReleasesDGX agent

arXiv:2608.02769v1 Announce Type: cross Abstract: Multimodal supervised learning seeks to leverage multiple heterogeneous data sources to improve predictive performance. A central challenge is determi

DataSpace: Benchmarking Data Agents for Verifiable Analytics over Heterogeneous Workspaces

Model ReleasesDGX agent

arXiv:2608.03451v1 Announce Type: new Abstract: Data agents enable natural-language analytics over organizational workspaces, where relevant evidence may be scattered across databases, structured file

Decoupling Generation and Selection for Budget-Constrained Faithful Summarization

ResearchDGX agent

arXiv:2608.03655v1 Announce Type: cross Abstract: Abstractive summarization models remain vulnerable to factual inconsistency, redundancy, and weak length control. We propose a modular generation-and-

Deep Divide-and-Reduce in Symbolic Regression

ResearchDGX agent

arXiv:2608.02628v1 Announce Type: cross Abstract: Symbolic regression (SR) is the task of discovering underlying patterns from data and representing them using mathematical expressions. Current machin

Deep Frequency-Aware Functional Maps for Robust Shape Matching

ResearchDGX agent

arXiv:2402.03904v3 Announce Type: replace Abstract: Deep functional map frameworks are widely employed for 3D shape matching. However, most existing deep functional map methods cannot adaptively captu

DeepSeek V4 Flash 0731 at 10–17 t/s (nothink) on MacBook M5 Pro **64GB***, partly via SSD streaming

Model ReleasesDGX agent

Inspired by a post from u/giveen I motivated claude (no patinence on my side to work through everything myself) to help me get DS running on my MacBook M5 Pro 64GB and it exceeded my expectations.. be

DeepSeek V4 Flash 0731 is ready to fine-tune on Fireworks. Run SFT, DPO, and RL training jobs on the Dedicated Training API. SFT and DPO fro…

Model ReleasesDGX agent

DeepSeek V4 Flash 0731 is ready to fine-tune on Fireworks. Run SFT, DPO, and RL training jobs on the Dedicated Training API. SFT and DPO from the managed UI. Built for coding agents and high-volume pr

DeepSeek V4 Flash is Ollama's fastest growing model ever in token usage, and the most popular model on OpenRouter this week. It’s available …

Model ReleasesDGX agent

DeepSeek V4 Flash is Ollama's fastest growing model ever in token usage, and the most popular model on OpenRouter this week. It’s available in Pi across a range of providers. If you’ve never tried an

Deepseek V4 Flash just hit Colibri, does anyone have numbers?

Model ReleasesDGX agent

I'm mosty interested in 128-192GB VRAM with 128-256GB RAM to spare, so SSD streaming is basically not even necessary. Seems only FP4 is supported, so older hardware will likely be slow - no Unsloth GG

DeepSeek-V4-Flash on SM89 4x48gb 4090s with DSpark

Model ReleasesDGX agent

https://github.com/yhfgyyf/vllm-deepseek-v4-sm89 I couldn't believe that someone actually got vLLM working with this particular set of GPUs, but here it is. The video is from right after I got it work

Deferred Exposure of Future Trajectories for Verifiable Reasoning in Autonomous Driving VLMs

SafetyDGX agent

arXiv:2608.01755v2 Announce Type: replace Abstract: Recent Vision-Language-Action (VLA) models for autonomous driving (AD) increasingly utilize chain-of-thought (CoT) supervision to enhance the reason

Demis Hassabis is stepping down as CEO of Google DeepMind to be the unit's chairman and will add the title of Alphabet chief scientist; GOOG drops 3%+ (Axios)

IndustryDGX agent

Axios: Demis Hassabis is stepping down as CEO of Google DeepMind to be the unit's chairman and will add the title of Alphabet chief scientist; GOOG drops 3%+ — Demis Hassabis is leaving his role as CE

DenialRAG: Single-Document RAG Poisoning via Embedded Parametric Denial

Model ReleasesDGX agent

arXiv:2608.02678v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) systems are vulnerable to corpus poisoning: an attacker who inserts a crafted document into the retrieval corpus

DeRP: An Algorithm for Self-Assembly of Power-Delivery Networks using Recursive Branching in Information-Limited Environments

ResearchDGX agent

arXiv:2608.02904v1 Announce Type: new Abstract: Delivering sustained power to distributed equipment in unstructured field environments using pre-planned wired networks or battery-based solutions prese

Design and Evaluation of an AI-Enabled Cloud-Edge Architecture for Connected Precision Agriculture Farms

AgentsDGX agent

arXiv:2608.03816v1 Announce Type: new Abstract: Plant diseases cause significant yield losses worldwide, with tomato crops particularly susceptible to early blight, late blight, and leaf mold. Manual

Design Criteria for SGD Preconditioners: Local Conditioning, Noise Floors, and Basin Stability

ResearchDGX agent

arXiv:2511.19716v3 Announce Type: replace-cross Abstract: Stochastic Gradient Descent (SGD) often slows in the late stage of training due to anisotropic curvature and gradient noise. We analyze precon

Design-Time Optimization of Deep Neural Networks for Intermittent Learning on Microcontrollers

Local AiDGX agent

arXiv:2608.03589v1 Announce Type: new Abstract: We present a method for designing deep neural networks (DNNs) for intermittent, energy-autonomous, on-device learning on microcontroller units (MCUs). I

Designing a Good Virtual Node: Addressable and Cardinality-Preserving Global Memory for Message Passing Architectures

ResearchDGX agent

arXiv:2608.02709v1 Announce Type: cross Abstract: Virtual nodes give message-passing neural networks a simple global communication route, but the standard node--VN--node pipeline compresses the graph

Designing Social Robots for Inclusive Child Wellbeing Assessment: Insights from Communities Supporting Developmental Language Disorder and Forced Migration

ResearchDGX agent

arXiv:2608.03820v1 Announce Type: new Abstract: Assessing children's wellbeing and mental health can be particularly challenging for children experiencing communication barriers, such as children with

Detecting Hallucinations and Recovering Verified Answers in Arabic Islamic Question Answering

Model ReleasesDGX agent

arXiv:2608.03720v1 Announce Type: new Abstract: Large language models can generate fluent responses to Islamic questions while introducing factual errors that are difficult to identify. This paper pre

Detecting high-frequency brain disorder signals using dynamic mode decomposition from EEG

ResearchDGX agent

arXiv:2608.02804v1 Announce Type: cross Abstract: Recent studies have reported clearly identifiable dynamical changes in the high-frequency range of EEG signals recorded during specific stimuli, such

Detecting Pose Estimation Failures via Keypoint Self-Consistency

ResearchDGX agent

arXiv:2608.03516v1 Announce Type: new Abstract: One common approach to pose estimation involves predicting object keypoints in an image, followed by using Perspective-n-Point algorithms to compute the

Developers in Africa are increasingly choosing Chinese open-source AI models over US models, saying they are downloadable, easier to customize, and much cheaper (New York Times)

IndustryDGX agent

New York Times: Developers in Africa are increasingly choosing Chinese open-source AI models over US models, saying they are downloadable, easier to customize, and much cheaper — Developers built Sunf

DiagChain: A Diagnostic Benchmark for Evaluating LLM Agents on Evidence-Grounded Attack Chain Reconstruction

Model ReleasesDGX agent

arXiv:2608.03591v1 Announce Type: cross Abstract: Large Language Model (LLM) agents offer a promising approach to attack chain reconstruction by retrieving and interpreting heterogeneous telemetry to

DiagLoop: A Counterfactual Data Flywheel with Stage-Localized Reinforcement for Diagnostic LLMs

Local AiDGX agent

arXiv:2608.03674v1 Announce Type: new Abstract: Causal diagnostic models must explain how conclusions follow from evidence because diagnoses guide repairs and treatments. Yet serious cases are scarce,

DiffImaginE: Imagine to Verify Entity Types with Diffusio

TutorialsDGX agent

arXiv:2608.03025v1 Announce Type: new Abstract: Multimodal named entity recognition (MNER) determines whether each candidate span and entity-type hypothesis is supported by joint textual and visual ev

DigitCode: Symbolic Tokenization of Hand Motion by Anatomical Units

ResearchDGX agent

arXiv:2608.03127v1 Announce Type: cross Abstract: Hand motion carries the finest-grained information in human activity, yet the representations behind hand generation, understanding, and robot learnin

Disentangling Language Modeling and Boundaries

SafetyDGX agent

arXiv:2608.03599v1 Announce Type: new Abstract: Byte-level language models are usually argued for on the grounds of robustness, multilingual fairness, and character-level skills. We point to a differe

Disentangling MLP Neuron Weights in Vocabulary Space

Model ReleasesDGX agent

arXiv:2604.06005v2 Announce Type: replace Abstract: Interpreting the information encoded in language model weights remains a fundamental challenge in mechanistic interpretability. In this work, we int

Distilled Roads: Generalisable Road Network Extraction Across Sensors, Resolutions, and Region

ResearchDGX agent

arXiv:2608.03407v1 Announce Type: cross Abstract: Road network segmentation from satellite imagery remains challenging due to large geographic variation in road appearance, occlusions, and domain shif

Distractor-Aware Truncation: Disentangling Context-Length Effects from Signal Loss in Long-Context LLM Benchmarks

Model ReleasesDGX agent

arXiv:2608.03297v1 Announce Type: new Abstract: A standard claim in the literature on retrieval-augmented and memory-augmented language models is that shorter context is better when the relevant infor

DiverseDiT++: Quantifying, Analyzing, and Promoting Representation Diversity in Diffusion Transformers

SafetyDGX agent

arXiv:2608.03082v1 Announce Type: new Abstract: Recent advances in Diffusion Transformers (DiTs) have enabled remarkable progress in visual synthesis, benefiting from their superior scalability. To fa

Diversity is Not Ambiguity: Toward Accurate and Efficient Ambiguity Detection for Open-Domain QA

Model ReleasesDGX agent

arXiv:2608.03177v1 Announce Type: new Abstract: How can question answering (QA) systems determine whether a query is ambiguous? Ambiguity detection is essential in open-domain QA, as misclassification

Divide-and-Conquer: Towards Generalizable Amortized Bayesian Inference for the Drift Diffusion Model

TutorialsDGX agent

arXiv:2608.03566v1 Announce Type: cross Abstract: The drift diffusion model (DDM) is a cornerstone of cognitive decision-making research. Although numerous estimation methods exist, researchers contin

DocTrace: Towards Traceable Long Document VQA via Hierarchical Evidence Graph Reasoning

SafetyDGX agent

arXiv:2608.03292v1 Announce Type: new Abstract: Long Document Visual Question Answering (LongDocVQA) requires Multimodal Large Language Models (MLLMs) to locate, integrate, and reason over heterogeneo

Document OCR is not Getting Commoditized (by Frontier Models) The most common question I get is whether frontier models are going to eat all…

Model ReleasesDGX agent

Document OCR is not Getting Commoditized (by Frontier Models) The most common question I get is whether frontier models are going to eat all document processing solutions - just screenshot the page an

Does Forgetting Transfer Across Modalities? A Real-World Benchmark for Cross-Modal Knowledge Unlearning Evaluation

Model ReleasesDGX agent

arXiv:2608.03791v1 Announce Type: new Abstract: Vision-Language Models (VLMs), like Large Language Models (LLMs), may memorize sensitive, copyrighted, or harmful knowledge from their pretraining corpo

Don't Let Me Ask for It: LLMs Show Deficiencies in Active Multi-Turn Information Acquisition for Abductive Inference

ResearchDGX agent

arXiv:2608.03388v1 Announce Type: new Abstract: Abductive reasoning requires forming hypotheses that explain observed evidence and revising them as new evidence becomes available. While large language

Don't Peek at the Answer: Outcome-Masked Group Relative Policy Optimization for Label-Free RLVR

SafetyDGX agent

arXiv:2608.03119v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) improves LLM reasoning but typically relies on ground-truth (GT) answers, limiting scalability. Vo

Don't Regenerate, Debug: A Domain-Specific Agent for Repairing Near-Miss Hardware Operators

AgentsDGX agent

arXiv:2608.02712v1 Announce Type: cross Abstract: Kernel generation for hardware accelerators such as GPUs and NPUs has become a proving ground for large language models (LLMs), and state-of-the-art s

Don't Walk the Line: Boundary Guidance for Filtered Generation

Model ReleasesDGX agent

arXiv:2510.11834v3 Announce Type: replace-cross Abstract: Generative models are increasingly paired with safety classifiers that filter harmful or undesirable outputs. A common strategy is to fine-tun

dots.tts.edit: Precisely Controlled Speech Editing with a Continuous Autoregressive Model

Model ReleasesDGX agent

arXiv:2608.02673v1 Announce Type: cross Abstract: Speech editing for content creation requires precise control over both what an edit should do and where it should apply. Free-form natural language pr

Double Descent in Gradient Boosting Decision Trees via Split-Candidate Scaling

Model ReleasesDGX agent

arXiv:2608.03111v1 Announce Type: new Abstract: Double descent is commonly studied by scaling an explicit capacity parameter, such as neural-network width. For gradient boosting decision trees (GBDTs)

Double Down on Defense: Strengthening Deep Perceptual Hashes against Evasion Attacks without Retraining

SafetyDGX agent

arXiv:2608.03101v1 Announce Type: new Abstract: Near-duplicate image matching is crucial for trust and safety, provenance verification, copyright enforcement, and large-scale visual search. Modern pla

DP-MemView: A Memory Interface for Attribute-Level Transcript Privacy in Long-Term LLM Agents

Model ReleasesDGX agent

arXiv:2608.03130v1 Announce Type: cross Abstract: Long-term memory enables persistent personalization in LLM agents, but repeated memory-conditioned responses can cumulatively reveal protected attribu

Dr. AGENTONOMICS: A Didactic Experiment of AGENTONOMICS

AgentsDGX agent

arXiv:2608.03524v1 Announce Type: new Abstract: AGENTONOMICS is a framework that treats AI agents as economic entities that can be designed, managed, and governed through an integrated management arch

← Previous
1…9495969798…1408
Next →