AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
84,460 results
6 Aug 2026

Training-Free Hashing-Based Attention via Binary Principal Components

Local AiDGX agent

arXiv:2608.04405v1 Announce Type: cross Abstract: Long-context large language models (LLMs) are increasingly deployed in real-world applications, yet self-attention remains a major efficiency bottlene

Transfer Learning for Named Entity Recognition of Classical Latin through LLM Prompting

Model ReleasesDGX agent

arXiv:2608.04015v1 Announce Type: new Abstract: With the increase in digitized resources of Classical Latin texts and modern breakthroughs of Large Language Models (LLMs), I contribute to ancient lang

Transferable Dual-Stream Representations for Mesoscale-Preserving Sea Surface Temperature Downscaling

ResearchDGX agent

arXiv:2608.04230v1 Announce Type: cross Abstract: Deep learning models for scientific spatio-temporal downscaling often minimize reconstruction error while failing to preserve physically meaningful mu

Content type
AllBlogX PostPaperYouTubeRedditGitHub

TRCoRSurg: Temporal-Relational Co-Reasoning for Surgical Video Triplet Recognition

ResearchDGX agent

arXiv:2608.04606v1 Announce Type: new Abstract: Understanding complex surgical scenes requires recognizing multiple interdependent entities, such as instruments, actions, and targets, while maintainin

TriCLE: Tri-Modal Vision-Language Reasoning for Edge-Deployed Fine-Grained Clustering

SafetyDGX agent

arXiv:2608.04175v1 Announce Type: new Abstract: Edge platforms used for aerial observation must interpret aircraft imagery under limited memory, limited compute, and intermittent connectivity. This se

Trident : How to Break Deep Reinforcement Learning Cyber Defenses (Agentic)

Model ReleasesDGX agent

arXiv:2608.04317v1 Announce Type: cross Abstract: Autonomous cyber defense systems based on Deep Reinforcement Learning (DRL) have attracted significant research attention, yet remain evaluated almost

TRNet: Topography-Guided Frequency Rectification and Structure-Aware Decoding for Multimodal Paddy Rice Segmentation

ResearchDGX agent

arXiv:2608.04154v1 Announce Type: cross Abstract: Mapping paddy rice from very-high-resolution imagery in mountainous and hilly regions is difficult because terrain alters optical appearance and incre

Tropical Algebraic Geometry for Neuronal Representations: An Arakelov-Green Measure Based Descriptor for Graph Learning

Model ReleasesDGX agent

arXiv:2608.04460v1 Announce Type: cross Abstract: The quantitative analysis of 3D neuronal morphologies requires capturing both graph topology and spatial geometry. Current message-passing Graph Neura

TS2TabPFN: Time Series Classification and Extrinsic Regression through Feature Extraction and a Tabular Foundation Model

ResearchDGX agent

arXiv:2608.04174v1 Announce Type: new Abstract: Time series data are ubiquitous in practical applications, where classification (TSC) and extrinsic regression (TSER) have emerged as essential tasks fo

TwinIR: Coordinated Invisible Dual-Point Attacks on Online HD Map Construction

AgentsDGX agent

arXiv:2608.04453v1 Announce Type: cross Abstract: Online HD map construction is critical to prediction and planning in autonomous driving. We find that existing physical attacks against online map con

Two-Level Meta-Rubrics for Evaluating Open-Ended Generation: GAMUT, a Benchmark for Factual Completeness

Model ReleasesDGX agent

arXiv:2607.19322v2 Announce Type: replace Abstract: Rubric-based evaluation of open-ended generation faces a fundamental tension between expressiveness and reliability. Authoring a faithful rubric req

Uber burned through its 2026 AI coding budget in four months. Microsoft canceled most of its Claude Code licenses six months after rolling t…

Model ReleasesDGX agent

Uber burned through its 2026 AI coding budget in four months. Microsoft canceled most of its Claude Code licenses six months after rolling them out. The mechanics are simple: per-token cost keeps fall

UBLLIE: Unified Backlight and Low-Light Image Enhancement

ApplicationsDGX agent

arXiv:2608.04429v1 Announce Type: new Abstract: Backlit and low-light images often suffer from severe exposure imbalance or global underexposure, presenting significant challenges for both visual perc

UG-UMRE: Uncertainty-Guided Modality Augmentation and Distributional Calibration for Unified Multimodal Relation Extraction

Model ReleasesDGX agent

arXiv:2608.04949v1 Announce Type: cross Abstract: Unified Multimodal Relation Extraction (UMRE) aims to identify intra-modal and cross-modal relations between textual entities and visual objects. Howe

Uncertainty-aware Predict-Then-Optimize Framework for Equitable Post-Disaster Power Restoration

ResearchDGX agent

arXiv:2508.04780v2 Announce Type: replace-cross Abstract: The increasing frequency of extreme weather events, such as hurricanes, highlights the urgent need for efficient and equitable power system re

Understanding Fault Tolerance of Adversarially Robust Pruned Models

ResearchDGX agent

arXiv:2608.04173v1 Announce Type: new Abstract: Deep neural networks (DNNs) deployed on resource-constrained neuromorphic hardware face three concurrent challenges: the need for model compression thro

Unforgettable Generalization in Language Models

TutorialsDGX agent

arXiv:2409.02228v2 Announce Type: replace-cross Abstract: When language models (LMs) are trained to forget (or 'unlearn'') a skill, how precisely does their behavior change? We study the behavior of t

Unifying quantum measurement constructions via a relative-entropy minimum change principle

ResearchDGX agent

arXiv:2608.04055v1 Announce Type: cross Abstract: The minimum change principle provides an information-theoretic characterization of the Bayes reversal channel in classical probability theory and has

UniWorld-View: Large-Baseline View Synthesis via Video Diffusion Models

Model ReleasesDGX agent

arXiv:2608.04701v1 Announce Type: new Abstract: The abundance of casually captured monocular videos and images on social media provides a valuable source for immersive content creation, where generati

Unleashing the Potential of Vision-Language Models for Generalizable AI-Generated Image Detection

ResearchDGX agent

arXiv:2608.04935v1 Announce Type: new Abstract: Recent work has shown that a simple linear probe on frozen representations from modern vision foundation models (VFMs) can achieve state-of-the-art AIGI

Unscented KalmanNet: a hybrid deep learning filter with calibrated posterior covariance for nonlinear state estimation

SafetyDGX agent

arXiv:2608.04201v1 Announce Type: new Abstract: State estimation for nonlinear dynamical systems is commonly performed with the Unscented Kalman filter (UKF), which propagates the state moments throug

Unsloth's Gemma 4 mmproj silently broke vision & audio on newer llama.cpp builds — anyone else hit this?

Model ReleasesDGX agent

So I had been building ScreenMind, kinda like local ai desktop assistant that uses Gemma 4 for screen analysis, voice memo transcription, and meeting transcription — all through llama-server. Everythi

Variational Bounds for Perceptron Learning from Structured Data

ResearchDGX agent

arXiv:2608.04882v1 Announce Type: new Abstract: We introduce a variational approach to a finite-temperature continuous-spin perceptron trained on a Gaussian mixture. The model allows for a broad class

Visual Anchoring in Diffusion: Multimodal Zero-Shot Skeleton Action Recognition

ResearchDGX agent

arXiv:2608.04623v1 Announce Type: new Abstract: Zero-shot Skeleton Action Recognition (ZSAR) remains ambiguous when unseen actions share similar skeleton joint dynamics but differ in objects or scene

Visual Representation Matters: Exploiting Temporal Differences in Video-to-Audio Generation

Model ReleasesDGX agent

arXiv:2608.04902v1 Announce Type: new Abstract: Video-to-audio (V2A) generation extends image-to-audio generation (I2A) by introducing consecutive frames that provide essential temporal cues for audio

Visualizing Graph-to-Answer Mechanism Recovery in Materials-Science Hypothesis Generation

ApplicationsDGX agent

arXiv:2608.04170v1 Announce Type: cross Abstract: AI co-scientists can generate fluent materials-science hypotheses, but fluency does not show that an answer preserves a scientifically meaningful mech

VoxStruct3D: Structure-Leading Flow Matching for Voxel-Space 3D MRI Synthesis

SafetyDGX agent

arXiv:2608.04557v1 Announce Type: new Abstract: High-fidelity 3D MRI synthesis requires both globally coherent anatomy and fine-grained voxel-level detail. Although latent diffusion makes volumetric g

VQ-VAD: Vector-quantized Motion Representation Learning for Human-centric Video Anomaly Detection

TutorialsDGX agent

arXiv:2608.05069v1 Announce Type: cross Abstract: Video Anomaly Detection (VAD) is inherently challenging due to the scarcity of anomalies and the large visual variability in surveillance footage, inc

War in the Abstract: The Rise and Consequences of Militarized Language in Scientific Communication

SafetyDGX agent

arXiv:2606.23462v2 Announce Type: replace-cross Abstract: Scientists do not, by profession, wage war. Yet warfare's vocabulary consistently appears in their abstracts. To quantify the extent to which

We made an MCP for your phone. Your laptop is just half of your life, and the other half is in your phone. Now your agent gets the mobile sc…

Model ReleasesDGX agent

We made an MCP for your phone. Your laptop is just half of your life, and the other half is in your phone. Now your agent gets the mobile screen too. No connectors, no complex setup, it has access to

We re-tested GPT-5.6 Luna from @OpenAI on ARC-AGI (Verified) following its recent 80% price reduction: - ARC-AGI-2: 59.6%, $0.18/task - ARC-…

Model ReleasesDGX agent

We re-tested GPT-5.6 Luna from @OpenAI on ARC-AGI (Verified) following its recent 80% price reduction: - ARC-AGI-2: 59.6%, 0.18/task - ARC-AGI-1: 90.7%, 0.07/task The new results match Luna's original

We're expanding our partnership with @huggingface to accelerate open science. Our storage on the Hub is roughly tripling to ~2 petabytes, & …

IndustryDGX agent

We're expanding our partnership with @huggingface to accelerate open science. Our storage on the Hub is roughly tripling to ~2 petabytes, & our downloads now run at high speed—even for our largest dat

We’re making better intelligence easier to access in ChatGPT for everyone: - GPT-5.6 Sol now powers both Instant and deep reasoning for Plus…

Model ReleasesDGX agent

We’re making better intelligence easier to access in ChatGPT for everyone: - GPT-5.6 Sol now powers both Instant and deep reasoning for Plus & Pro users, delivering more factual, focused responses. -

What are Agentic Workflows?

AgentsDGX agent

**Agentic Workflows** enable automated systems to execute complex tasks by orchestrating multiple agents that interact with each other and external services. They delegate subtasks, monitor progress,

What Is a Skill Worth? Structure-Aware Shapley Valuation of Agent Skills

AgentsDGX agent

arXiv:2608.04562v1 Announce Type: new Abstract: Agent skills are increasingly optimized by automated feedback loops, producing long structured artifacts whose internal value remains unclear. We study

What We Observe as LLM Behavior Can Be a Side-effect of Inference Backend

Model ReleasesDGX agent

arXiv:2608.04714v1 Announce Type: cross Abstract: Benchmark scores are reported as properties of a model, yet the inference framework used to produce them, such as HuggingFace, vLLM, or Ollama, are co

When Absence Is Evidence: Evaluating Completeness-Sensitive Negative Reasoning in Large Language Models

ResearchDGX agent

arXiv:2608.04591v1 Announce Type: cross Abstract: Large language models (LLMs) are often asked whether something is absent from a record, list, or retrieved context. Yet non-observation licenses a neg

When Diffusion Models Forget Who You Are: Identity Preservation in Face Inpainting under Large Occlusions

TutorialsDGX agent

arXiv:2608.04820v1 Announce Type: new Abstract: Face inpainting with diffusion models has recently achieved impressive visual quality, yet preserving identity fidelity under significant occlusion and

When Does Latent Communication Pay? A Causal Audit of Relayed KV Caches in Multi-Agent LLMs

Model ReleasesDGX agent

arXiv:2608.04893v1 Announce Type: cross Abstract: Multi-agent LLM systems relay key--value caches instead of text and credit their gains to exchanged ``latent thoughts''. That credit is a claim about

When does training on downscaled images yield the same gradients?

ResearchDGX agent

arXiv:2608.04448v1 Announce Type: cross Abstract: Diffusion transformers deliver strong image generation, but their training cost grows superlinearly with resolution. Recent work justifies training or

When Large Language Models Know the Table: A Framework for Assessing Data Contamination in Tabular Datasets

ResearchDGX agent

arXiv:2510.20351v3 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly exposed to data contamination, i.e., performance gains driven by prior exposure of test datasets

When Memory Lies: An Empirical Study of Spatial Memory Staleness in VLM Agents

SafetyDGX agent

arXiv:2608.04574v1 Announce Type: new Abstract: Memory-augmented VLM agents act on persistent spatial knowledge, yet that knowledge silently goes stale as the environment changes. We ask what happens

When Modalities Fail to Tango: Conformal Backdoor Detection in Multimodal Contrastive Learning

ResearchDGX agent

arXiv:2608.04052v1 Announce Type: cross Abstract: Backdoor attacks in multimodal contrastive learning (MCL) have garnered growing attention in recent years, as many downstream tasks critically depend

When Modalities Remember: Continual Learning for Multimodal Knowledge Graphs

SafetyDGX agent

arXiv:2604.02778v2 Announce Type: replace Abstract: Real-world multimodal knowledge graphs (MMKGs) are dynamic, with new entities, relations, and multimodal knowledge emerging over time. Existing cont

When More Becomes Less: Position-Dependent Repetition Effects in Language Models

ResearchDGX agent

arXiv:2608.04021v1 Announce Type: new Abstract: Cloze-style probes that vary how often a target token appears implicitly assume that more copies of a target affect prediction the same way regardless o

When Prompts Become Pixels: Prompt-Region Grounding for Multimodal Reasoning

Model ReleasesDGX agent

arXiv:2608.04726v1 Announce Type: new Abstract: Multimodal large language models increasingly reason over screenshots and documents where the task itself may be written in pixels. Yet benchmarks usual

When Proxy Prediction Becomes Equation Reconstruction: Diagnostics and Residual Learning for Factor-Derived Proxy Supervision

ResearchDGX agent

arXiv:2608.04393v1 Announce Type: new Abstract: Scientific machine learning often relies on proxy targets computed from known domain factors when direct observations are limited. When those same facto

When Shared Rollouts Fail in Defensive Driving Evaluation: A NAVSIM Score Basis Audit

AgentsDGX agent

arXiv:2608.04896v1 Announce Type: new Abstract: Defensive driving scores are useful only when they preserve distinctions between policies that observe surrounding actors and those that do not. Re-simu

Why Ranking Anomaly Detection Algorithms Isn't as Reliable as You May Think

Model ReleasesDGX agent

arXiv:2608.04613v1 Announce Type: new Abstract: Anomaly detection is a safety-critical machine learning problem with applications ranging from fraud detection to network intrusion prevention and indus

WorldCycle: Self-Verifiable Reinforcement Learning for Long-Horizon Video World Models

Model ReleasesDGX agent

arXiv:2608.04964v1 Announce Type: new Abstract: Interactive video world models are essential for long-horizon planning and exploration, yet they suffer from compounding errors. Post-training methods s

Yesterday, my OpenAI collaborator and I gave a detailed talk on the Huggingface incident, our models creating 'the message board', model mis…

IndustryDGX agent

Yesterday, my OpenAI collaborator and I gave a detailed talk on the Huggingface incident, our models creating 'the message board', model misalignment, and more. https://www.youtube.com/watch?v=87DyyMV

YOLO-PVC: 2D-to-3D Consolidation of Slice-wise Detections for Volumetric Liver Tumor Localization in MRI

ResearchDGX agent

arXiv:2608.04642v1 Announce Type: new Abstract: Slice-wise 2D object detectors are increasingly applied to volumetric data due to their computational efficiency and scalability, yet they often yield f

YOLOv14:Unified Cross-Domain Real-Time Object Detectionwith Adaptive Multi-View Representation

SafetyDGX agent

arXiv:2608.04720v1 Announce Type: new Abstract: Real-time object detectors achieve remarkable accuracy under controlled conditions, yet degrade sharply on non-ideal inputs: fisheye distortion, game-re

Your agentic summer: No-cost lessons from Google experts to build and scale agents

Model ReleasesDGX agent

I’ve talked to developers, IT leaders, and builders who all ask the same question: How do we actually get agents into production? The answer isn't theoretical — it's hands-on. Whether it’s designing a

YouTube-Occ: Learning Indoor 3D Semantic Occupancy Prediction from YouTube Videos

SafetyDGX agent

arXiv:2506.18266v2 Announce Type: replace Abstract: 3D semantic occupancy prediction is crucial for fine-grained scene understanding, yet its advancement in privacy-sensitive indoor environments is fu

Zero-shot reasoning for simulating scholarly peer-review

Model ReleasesDGX agent

arXiv:2510.02027v2 Announce Type: replace Abstract: Scholarly publishing requires scalable scrutiny supported by auditable evidence. This paper presents a two-component benchmark of xPeer, the peer-re

Zero-shot Sim2Real Transfer for Magnet-Based Tactile Sensor on Insertion Tasks

TutorialsDGX agent

arXiv:2505.02915v2 Announce Type: replace Abstract: Tactile sensing is an important sensing modality for robot manipulation. Among different types of tactile sensors, magnet-based sensors, like u-skin

ZoomV: Temporal Zoom-in for Efficient Long Video Understanding

AgentsDGX agent

arXiv:2504.01407v3 Announce Type: replace-cross Abstract: Long video understanding poses a fundamental challenge for large video-language models (LVLMs) due to the overwhelming number of frames and th

5 Aug 2026

3DGSI-Assessor: A Large-Scale Dataset and An LMM-based Method for 3D Gaussian Splatting Image Quality Assessment

Model ReleasesDGX agent

arXiv:2608.03279v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) has become a dominant representation for real-time novel view synthesis (NVS), yet its storage footprint makes compression

40% speedup of MoE training with faster megakernel, by cursor, of all people (for B200s)

Local AiDGX agent

daily reminder not to trust benchmarks and run it yourself. claimed e2e speedup is ~40%, forwards are ~140% faster I would wager that compared to a naive kernel anyone can write it's more in the range

← Previous
1…9091929394…1408
Next →