AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

86,428Total entries
1Added by human
86,427Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,016 results
11 Aug 2026

TransNRank: Towards Accurate Neoantigen Ranking with Transformer

Local AiDGX agent

arXiv:2608.01924v2 Announce Type: replace-cross Abstract: Personalized neoantigen prediction is challenging due to the scarcity of positive samples, the noise of the experimental data, the severe clas

Uncertainty-Aware Variational Reward Factorization via Probabilistic Preference Bases for LLM Personalization

SafetyDGX agent

arXiv:2604.00997v2 Announce Type: replace Abstract: Reward factorization personalizes large language models (LLMs) by decomposing rewards into shared basis functions and user-specific weights. Yet, ex

VisionSelector: End-to-End Learnable Visual Token Compression for Efficient Multimodal LLMs

ResearchDGX agent

arXiv:2510.16598v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) encounter significant computational and memory bottlenecks from the massive number of visual tokens generat

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Weather- and Location-Aware Agentic Dining Recommendation: Leveraging LLM World Knowledge for Region-Sensitive Contextual Reasoning

Local AiDGX agent

arXiv:2608.07593v1 Announce Type: cross Abstract: Context-aware recommender systems have long recognized that factors such as location, time, and weather shape where and what people choose to eat. Exi

What Keeps Agent Skills from Being Reusable? Evidence from 138K SKILL.md Files

SafetyDGX agent

arXiv:2608.08453v1 Announce Type: new Abstract: Under the current standard, Agent Skills are SKILL.md files that combine instructions with supporting resources, enabling Large Language Model (LLM) age

10 Aug 2026

A Practical Evaluation Method for Long-Form Simultaneous Speech-to-Speech Translation

SafetyDGX agent

arXiv:2606.15059v2 Announce Type: replace Abstract: Simultaneous speech-to-speech translation (SimulS2ST) enables real-time cross-lingual communication, but existing evaluation has focused largely on

Autonomous discovery of accelerator commissioning algorithms

AgentsDGX agent

arXiv:2608.07138v1 Announce Type: cross Abstract: Simulated commissioning has become essential for de-risking modern light-source design and commissioning, but the procedures being simulated are still

Autonomy-of-Heads: Data-Free Sparse Attention from Frozen Query-Key Geometry

ApplicationsDGX agent

arXiv:2608.06849v1 Announce Type: cross Abstract: Long-context LLM inference is bottlenecked by quadratic attention computation and growing KV-cache costs. Existing sparse attention and KV-compression

Better Together: Quantifying the Benefits of AI-Assisted Recruitment

ResearchDGX agent

arXiv:2507.08029v2 Announce Type: replace Abstract: Hiring algorithms have mostly scored the materials recruiters already see. Large language models (LLMs) can instead generate new information about c

Beyond Attention: Signed Integrated Gradients Attribution in a BiomeGPT-Style Microbiome Transformer

ResearchDGX agent

arXiv:2608.06486v1 Announce Type: new Abstract: In a feature-tokenized transformer (arXiv:2106.11959) such as BiomeGPT (doi:10.64898/2026.01.05.697599), each input token is built by fusing a fixed ide

Beyond Co-Movement: Locality by Exposures Enables a Joint Factor-Graph Framework for Portfolio Diversification

ResearchDGX agent

arXiv:2608.06618v1 Announce Type: cross Abstract: Current portfolio construction methods are either agnostic to the effects of idiosyncratic shocks (standard factor models) or to the latent data struc

Beyond Visibility: Real-Time Surface Accessibility Fields from Sparse LiDAR

HardwareDGX agent

arXiv:2608.06412v1 Announce Type: cross Abstract: Understanding which surfaces in a scene are physically accessible to a given tool is fundamental for robotic interaction, yet 3D perception systems ty

Can MLLMs Decode the Creative Leap? Introducing C4 for Cross-Concept Understanding

ApplicationsDGX agent

arXiv:2608.06501v1 Announce Type: new Abstract: Creative capabilities of MLLMs matter in design, communication, education, and human--AI collaboration, yet remain difficult to evaluate because explici

Cloud-Boosted Low-Compute Multi-Channel Speech Enhancement

Local AiDGX agent

arXiv:2608.07423v1 Announce Type: cross Abstract: Low-latency, low-compute speech enhancement is essential for wearable devices with real-time communication requirements, but strict computational cons

CloudDiffusion: Diffusion-Based Scene Completion in the Point Cloud Domain

AgentsDGX agent

arXiv:2606.16048v3 Announce Type: replace Abstract: Reconstructing dense 3D scenes from sparse LiDAR point clouds (LiDAR scene completion) is a fundamental challenge in autonomous driving, where diffu

ConstructCIE: A Dataset for Extracting Causal Information from Construction Accident Narratives

ResearchDGX agent

arXiv:2608.06495v1 Announce Type: new Abstract: Construction accident narratives contain rich causal information, but the evidence is often implicit, long-span, and distributed. We introduce Construct

Corrupting Attention: Evasion-Based Adversarial Attacks on Encoder Attention in Detection Transformers

SafetyDGX agent

arXiv:2608.06674v1 Announce Type: new Abstract: Adversarial vulnerabilities remain a major concern for the safe deployment of neural networks, particularly in object detection, a core task embedded in

CrystalGRPO: Target-Aligned and Coverage-Preserving Reinforcement Learning for Flow-Based Crystal Structure Prediction

SafetyDGX agent

arXiv:2608.06582v1 Announce Type: new Abstract: Flow-based generative models can efficiently produce candidate structures for crystal structure prediction (CSP), but their pretrained objectives do not

Direct Visual Grounding by Directing Attention of Visual Tokens

ApplicationsDGX agent

arXiv:2511.12738v2 Announce Type: replace Abstract: Vision Language Models (VLMs) mix visual tokens and text tokens. A puzzling issue is the fact that visual tokens most related to the query receive l

DynaCrys: Crystal Generation with Dynamic Space-Group Diffusion

ResearchDGX agent

arXiv:2608.07401v1 Announce Type: cross Abstract: The search for new crystalline materials spans an enormous compositional and structural space. Generating candidates in this space requires jointly mo

Every Cache Entry Earns Its Place: Global Allocation of Resolution and Coverage for KV Cache Compression

Local AiDGX agent

arXiv:2608.07001v1 Announce Type: new Abstract: As large language models (LLMs) process increasingly long contexts, KV cache storage and repeated access have become a major bottleneck. Existing KV cac

Evolving Parallel Algorithm Portfolios via Potential-Aware Instance Generation with LLMs

Local AiDGX agent

arXiv:2608.06808v1 Announce Type: new Abstract: The Automatic Construction of Portfolios via Large Language Models (LLM-ACP) suffers from poor generalization in practical few-shot scenarios when solvi

Learning to Walk With Less: A Dyna-Style Approach to Quadrupedal Locomotion

SafetyDGX agent

arXiv:2509.06296v2 Announce Type: replace-cross Abstract: Traditional on-policy reinforcement learning (RL) controllers for quadrupedal locomotion often suffer from low data efficiency, requiring mill

Limit Points of Reflow with Minibatch Optimal Transport

TutorialsDGX agent

arXiv:2608.07042v1 Announce Type: cross Abstract: Rectified flows, also called flow matching or stochastic interpolants, are generative models that learn a time-dependent vector field steering a proba

Mind the Gap: A Dual Knowledge Graph Framework for Unified Multi-task User Intent Inference

SafetyDGX agent

arXiv:2608.06752v1 Announce Type: new Abstract: This paper proposes DKG-MTI, a dual knowledge graph framework for unified multi-task user intent inference from online travel reviews. Existing approach

Multimodal Drivers' Emotion Recognition and Safety-Oriented Intervention for Intelligent Transportation Systems

SafetyDGX agent

arXiv:2608.06378v1 Announce Type: cross Abstract: Driver emotions can affect risk perception, decision-making, and vehicle control under complex road conditions. Existing studies mainly focus on drive

MuST-VAD: Mutual Structured Learning for Video Anomaly Detection

TutorialsDGX agent

arXiv:2608.06913v1 Announce Type: new Abstract: In this paper, we propose MuST-VAD, a mutual structured learning framework for weakly supervised video anomaly detection (VAD) in which an anomaly detec

Online Conformal Prediction Beyond Feedback

SafetyDGX agent

arXiv:2608.07139v1 Announce Type: new Abstract: Uncertainty quantification is essential when deploying machine learning models in safety-critical applications. Online conformal prediction (OCP) provid

Optimizing Spectral Prediction in MXene-Based Metasurfaces Through Multi-Channel Spectral Refinement and Savitzky-Golay Smoothing

ResearchDGX agent

arXiv:2602.08406v2 Announce Type: replace-cross Abstract: The prediction of electromagnetic spectra for MXene-based solar absorbers, where MXenes are a family of two-dimensional transition metal carbi

Solver-Guided Reasoning for Mixed-Equilibrium Strategies

TutorialsDGX agent

arXiv:2608.06741v1 Announce Type: new Abstract: Reasoning in large language models (LLMs) is often grounded in human text, human demonstrations, and human-generated rationales. For equilibrium reasoni

Sources: officials say OpenAI risks its White House relationship by hiring Dean Ball, who has criticized Trump's AI strategy after leaving the administration (Thomas Barrabi/New York Post)

SafetyDGX agent

Thomas Barrabi / New York Post: Sources: officials say OpenAI risks its White House relationship by hiring Dean Ball, who has criticized Trump's AI strategy after leaving the administration — Top Open

SubtleTalk: Generating Controllable Weakly-correlated Facial Dynamics for 3D Talking Heads via Residual Flow Matching

ResearchDGX agent

arXiv:2608.06408v1 Announce Type: cross Abstract: Audio-driven 3D facial animation aims to synthesize realistic and temporally coherent motions from speech. Despite notable progress in lip synchroniza

Summarize First, Download Later: Onboard VLMs for Bandwidth-Efficient Earth Observation

HardwareDGX agent

arXiv:2608.06959v1 Announce Type: new Abstract: Modern Earth observation (EO) satellites carry increasingly advanced sensors that produce vast volumes of high-resolution, multispectral data, yet downl

TA-RAG: Tone Awareness as a Design Imperative for Retrieval-Augmented Generation

SafetyDGX agent

arXiv:2608.06672v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) has become a robust architecture for grounding large language models (LLMs) in trusted knowledge. However, standard

Vehicle routing problem using deep reinforcement learning - A case study about truck planning in the industry

AgentsDGX agent

arXiv:2608.06668v1 Announce Type: new Abstract: As an important component of the supply chain industry, transportation has experienced rapid development in the past decade with the assistance of digit

Vernata: Self-Supervised Learning of LiDAR Point Representations

ResearchDGX agent

arXiv:2608.06919v1 Announce Type: new Abstract: LiDAR serves as a primary sensing modality for robots operating in outdoor environments. However, the performance of deep learning models in this domain

When Coordination Becomes a Threat: Communication Attacks in LLM-Controlled Multi-Robot Systems

AgentsDGX agent

arXiv:2608.06830v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly used as high-level planners in embodied multi-robot systems, enabling robots to interpret natural language

9 Aug 2026

Okay hear me out new IQ test: see if an LLM gets smarter or dumb after fine-tuning on your stream of consciousness

AgentsDGX agent

Simon @disiok proposed a novel IQ test for language models, asking whether an LLM becomes smarter or dumber after being fine‑tuned on a user’s stream of consciousness. The tweet was posted on 9 Aug 20

8 Aug 2026

Building a zero-dependency C inference engine for BitNet (1.58-bit) - lessons from hitting 36 tok/s on a Xeon CPU

Local AiDGX agent

Over the past few months I have been building a CPU-first inference engine from scratch in pure C99 (no Python, no CUDA, no BLAS, just GCC and make). The focus has been running 1.58-bit ternary models

Imagine image 2.0, non-agentic yet, more to come in a week or two 💙

AgentsDGX agent

Imagine image 2.0, non-agentic yet, more to come in a week or two 💙 Announcing Imagine Image 2.0, our next generation image model with precision editing, crisp text rendering, improved factuality, and

7 Aug 2026

A Bridge from Audio to Video: Phoneme-Viseme Alignment Allows Every Face to Speak Multiple Languages

SafetyDGX agent

arXiv:2510.06612v2 Announce Type: replace Abstract: Speech-driven talking face synthesis (TFS) focuses on generating lifelike facial animations from speech input. Current TFS models perform well in En

A Reverse-BSDE Diffusion Sampler

ResearchDGX agent

arXiv:2505.06800v2 Announce Type: replace-cross Abstract: Diffusion-based generative models have renewed interest in stochastic differential equation methods for sampling from complex distributions. W

ASAT: Adaptive Scoring and Thresholding with Human Feedback for Robust Out-of-Distribution Detection

SafetyDGX agent

arXiv:2505.02299v2 Announce Type: replace-cross Abstract: Machine Learning (ML) models are trained on in-distribution (ID) data but often encounter out-of-distribution (OOD) inputs during deployment--

Beyond Frame Selection: Rethinking Long-Video Understanding with MLLMs

ResearchDGX agent

arXiv:2608.05592v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have achieved strong progress in video understanding, yet it remains challenging because the token limitation m

Beyond Sentiment: Comparing Traditional NLP and LLM-Based Multi-Dimensional Analysis for Political News Evaluation

SafetyDGX agent

arXiv:2608.05155v1 Announce Type: cross Abstract: Traditional sentiment analysis (SA) models, while effective for polarity classification, provide limited insight into the rhetorical, ideological, and

Bias Analysis of L2 Speaking Assessment Systems Using Concept Activation Vectors

SafetyDGX agent

arXiv:2608.06300v1 Announce Type: new Abstract: Automatic speaking assessment systems are increasingly deployed in high-stakes settings to mark second language (L2) learners' speaking tests, making it

CDSeg: A Renderable Gaussian Carrier for Image-to-3D Label Transfer

ResearchDGX agent

arXiv:2608.05482v1 Announce Type: new Abstract: Modern image models provide strong cues about what should be segmented in each view, but their masks do not by themselves determine where those labels s

CogVis: Must Open-Vocabulary Change Detection Perceive the Scene Anew for Every Query?

Local AiDGX agent

arXiv:2608.06150v1 Announce Type: new Abstract: Earth-surface monitoring requires change detection models capable of recognizing arbitrary semantic categories. Open-Vocabulary Change Detection (OVCD)

Coherence-Oriented Dream Scene Visualisation

ResearchDGX agent

arXiv:2608.05233v1 Announce Type: new Abstract: Dreams can be emotionally intense but difficult to communicate. We describe the Dream Scene Visualiser (DSV) system which turns written dream descriptio

Constraint-First Reasoning: A Training-Free Protocol for Exploiting Answer-Space Constraints in Mathematical Problem Solving

ResearchDGX agent

arXiv:2608.05254v1 Announce Type: new Abstract: Large language models can derive a plausible mathematical object yet still violate explicit requirements--for example, by omitting a modular reduction,

Discrete energy as an exact label-free training objective for finite-element surrogates

ResearchDGX agent

arXiv:2608.05437v1 Announce Type: cross Abstract: Supervised training of finite-element (FE) surrogate models requires reference solutions, and each reference solution is obtained by solving the syste

Does Latent Context Help? A Controlled Evaluation of Inverse Reinforcement Learning in Arctic Shipping

SafetyDGX agent

arXiv:2608.06105v1 Announce Type: cross Abstract: Artificial Intelligence (AI)-assisted navigation can help Arctic shipping adapt to rapidly changing sea-ice conditions, but reliable deployment requir

EdgeXpert: An Edge Device for Memory-Efficient LLM Inference with Mixture-of-Experts and Speculative Decoding

Local AiDGX agent

arXiv:2608.05303v1 Announce Type: cross Abstract: On-device deployment of Large Language Models (LLMs) has become essential for personalized edge applications. A primary bottleneck is external memory

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI

ResearchDGX agent

arXiv:2608.05258v1 Announce Type: new Abstract: Gradient-weighted Class Activation Mapping (Grad-CAM) is widely used to visualize model decisions, but it was originally formulated for convolutional ne

In the era of base LLM scaling (2022-2024), I believed the LLM line of research would reach a capability plateau (as later seen with base LL…

ResearchDGX agent

In the era of base LLM scaling (2022-2024), I believed the LLM line of research would reach a capability plateau (as later seen with base LLMs). In late 2024, after the o3 test-time compute demo, I ch

Learning visual representations for compositional analysis of artworks and photographs

SafetyDGX agent

arXiv:2608.06142v1 Announce Type: new Abstract: Composition, the deliberate arrangement of visual elements, is central to how meaning, emotion, and aesthetic quality are conveyed in artwork, yet it re

Mitigating Scoring Bias in LLM-as-a-Judge via Random Number Generation

SafetyDGX agent

arXiv:2608.05726v1 Announce Type: new Abstract: Large Language Models (LLMs) are often used as evaluators of text quality, known as LLM-as-a-Judge, which can outperform conventional automatic evaluati

Multi-Agent Reinforcement Learning for Online Traffic Scheduling in Time-Sensitive Application

SafetyDGX agent

arXiv:2608.05346v1 Announce Type: cross Abstract: Time-sensitive networking (TSN) is increasingly integrated into mobile edge computing (MEC) to support applications with stringent latency requirement

NeuroAdaptTrainer: A Fiji/ImageJ Plugin for YOLO-Based Neuron Segmentation, InteractiveCorrection and Transfer Learning

ResearchDGX agent

arXiv:2608.05226v1 Announce Type: new Abstract: Neuron counting and segmentation in microscopy images of neuronal cultures is a routine and time-consuming task in neuroscience research, traditionally

One Ranking, Any Budget: Matryoshka Evidence-to-Context Frame Selection for Long-Video Understanding

ResearchDGX agent

arXiv:2608.05707v1 Announce Type: new Abstract: Frame selection is essential for applying Large Multimodal Models (LMMs) to long videos due to severe frame redundancy and limited context windows. Sinc

← Previous
1…733734735736737…1034
Next →