AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
83,164 results
12 Aug 2026

The Multilingual Quantization Tax: Structural Collapse and Typological Fragility in Edge SLMs

Model ReleasesDGX agent

arXiv:2608.09941v1 Announce Type: new Abstract: While 4-bit weight quantization is critical for deploying Small Language Models (SLMs) on edge devices, evaluations of the resulting performance degrada

The Parser Already Knows: Lightweight Bias Correction in Constrained Decoding

SafetyDGX agent

arXiv:2608.10137v1 Announce Type: new Abstract: Grammar Constrained Decoding (GCD) forces Language Models (LMs) to produce syntactically valid outputs by masking out non-conforming tokens at each step

The Signal Rail: A Deterministic Motion Grammar for Communicating Conversational Agent State in Terminal Interfaces

AgentsDGX agent

arXiv:2608.10689v1 Announce Type: cross Abstract: Terminal interfaces to conversational agents report rich internal state (listening, thinking, executing tools, awaiting input, failing) almost entirel

Content type
AllBlogX PostPaperYouTubeRedditGitHub

The Truth Stays in the Family: Enhancing Contextual Grounding via Inherited Truthful Heads in Model Lineages

Model ReleasesDGX agent

arXiv:2606.15821v2 Announce Type: replace-cross Abstract: Recent advances in large language models (LLMs) have produced many specialized multimodal LLMs (MLLMs) that share common foundational LLMs, fo

ThinkAfford: Affordance-Centric Reasoning for Fine-Grained 3D Grounding in Cluttered Scenes

Local AiDGX agent

arXiv:2608.10981v1 Announce Type: new Abstract: Task-driven 3D affordance grounding aims to localize the functional region in a cluttered 3D scene that enables an action specified by a natural-languag

ThinkRetrieve: Retrieval-Augmented Reasoning Traces for Test-Time Scaling

TutorialsDGX agent

arXiv:2608.10928v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) improve performance by allocating additional inference-time compute to generate extended chain-of-thought reasoning. Howev

Threat-guided Policy-aware Scene Perturbation for Safe Autonomous Driving with Online Reinforcement Learning

SafetyDGX agent

arXiv:2608.10403v1 Announce Type: new Abstract: Reinforcement learning (RL) has shown promising performance in autonomous driving, yet ensuring the safety of online RL policies remains challenging due

Threshold Structure of Optimal Policies in Restart POMDPs

ResearchDGX agent

arXiv:2608.10936v1 Announce Type: cross Abstract: We study a Restart POMDP (Partially Observable Markov Decision Process) on a general Borel state space, where the controller either lets the hidden st

TideRL: Boosting Agentic RL Goodput with Readiness-Aware Scheduling

HardwareDGX agent

arXiv:2608.10402v1 Announce Type: new Abstract: Reinforcement learning (RL) for large language models is moving toward multi-turn agentic workloads, where rollout tasks repeatedly pause for external e

Tiered KV cache for large LLMs on Amazon SageMaker HyperPod with Curvine

HardwareDGX agent

Running large language model inference at scale forces a KV cache trade-off: oversized GPU instances or slow time-to-first-token. This post builds a tiered KV cache on Amazon SageMaker HyperPod that e

TimeRoute: Time-Aware Modality Routing and Diffusion for Multi-Modal Recommendation

ResearchDGX agent

arXiv:2608.10983v1 Announce Type: cross Abstract: Multi-modal recommenders fuse collaborative signals with item modalities such as text, images, and audio, but the usefulness of each drifts over time

Today, we’re adding another member to our model family. Meet North Micro Vision. Our smallest vision-language model yet, ideal for sophistic…

Model ReleasesDGX agent

Today, we’re adding another member to our model family. Meet North Micro Vision. Our smallest vision-language model yet, ideal for sophisticated document understanding. Available open-source under an

Together Serverless Inference gives developers a managed, high-throughput path for running Qwen3.8-2.4T-A95B across coding and agentic workl…

AgentsDGX agent

Together Serverless Inference gives developers a managed, high-throughput path for running Qwen3.8-2.4T-A95B across coding and agentic workloads. Start building: https://www.together.ai/models/qwen3-8

Token-Based Detection of Spurious Correlations in Vision Transformers

Local AiDGX agent

arXiv:2509.04009v2 Announce Type: replace-cross Abstract: Due to their powerful feature association capabilities, neural network-based computer vision models have the ability to detect and exploit uni

Topological Feasibility Guarantees for Differentiable Predictive Control

SafetyDGX agent

arXiv:2608.10332v1 Announce Type: cross Abstract: Differentiable predictive control (DPC), a self-supervised learning approach for approximating explicit model predictive control (MPC) policies, offer

Toward a Theory of Value in AI Alignment

SafetyDGX agent

arXiv:2608.10327v1 Announce Type: new Abstract: Can AI systems be aligned to human values? The popularization of large language models (LLMs) and multi-modal foundation models has seen a rise in harms

Toward Human Rights Benchmarking for LLMs: A Pilot Methodology

Model ReleasesDGX agent

arXiv:2608.10268v1 Announce Type: cross Abstract: Large language models (LLMs) increasingly mediate legal determinations over what human rights are realized, and how. Yet, no evaluation benchmark exis

Toward the Cognitive--Physical Limits of Embodied Intelligence through a World-Model-Centric Autonomous Racing Agent

Local AiDGX agent

arXiv:2608.10618v1 Announce Type: new Abstract: Embodied artificial intelligence aims to develop agents that perceive, reason, and act through continuous interaction with the physical world. However,

Towards Color-Faithful Low-Light Image Enhancement via Adaptive Color Debiasing and Saturation Rectification

SafetyDGX agent

arXiv:2608.10512v1 Announce Type: new Abstract: Low-light imaging often introduces color bias caused by the low signal-to-noise ratio and the image formation process. Although recent low-light image e

Towards Efficient Reasoning in LLM-Based Recommender Systems via Model Merging

Model ReleasesDGX agent

arXiv:2608.10447v1 Announce Type: cross Abstract: Large language model-based recommender systems are increasingly adopting slow-thinking models that generate step-by-step reasoning before making predi

Towards Geometry-Grounded Dense Semantic Matching with VGGT Priors

ResearchDGX agent

arXiv:2509.21263v2 Announce Type: replace Abstract: Semantic matching aims to establish pixel-level correspondences between instances of the same category and represents a fundamental task in computer

Towards Sustainable Artificial Intelligence: A Comprehensive Review and Comparative Analysis of Deep Learning Models' Carbon Footprint

ResearchDGX agent

arXiv:2608.09998v1 Announce Type: new Abstract: Artificial Intelligence (AI) and Machine Learning (ML) have become powerful tools for supporting and automating complex human tasks. Despite their benef

Towards Unified Dynamic Face Landmark Detection

Model ReleasesDGX agent

arXiv:2608.10346v1 Announce Type: cross Abstract: Although advancements in face landmark detection (FLD) methods continue to push performance boundaries, they overlook two major functional limitations

TrAC: Trace-Conditioned Answer Consistency for Efficient Uncertainty Quantification in LLMs

ResearchDGX agent

arXiv:2608.00422v2 Announce Type: replace Abstract: Large language models (LLMs) can generate fluent reasoning traces that nevertheless lead to incorrect answers, making response-level uncertainty est

TRACE-GS: On-Policy Trajectory Distillation with Privileged Geometric Conditioning for Sparse-View 3DGS Restoration

SafetyDGX agent

arXiv:2608.10286v1 Announce Type: new Abstract: We present TRACE-GS, an on-policy trajectory distillation framework that leverages privileged geometric conditioning at training time, thereby adapting

TRACE: Trustworthy Retrieval-Augmented Conversational Engine

Model ReleasesDGX agent

arXiv:2608.10176v1 Announce Type: new Abstract: Public service chatbots are expected to deliver recommendations from an underlying public service directory, while also making sure that the recommendat

Transformer Geometry Observatory TGO-IV: Developmental Topology Observatory

TutorialsDGX agent

arXiv:2608.09997v1 Announce Type: cross Abstract: Transformers have had a profound impact on the world of language processing and computer vision. As efforts to answer the million-dollar question of `

TransitReID: Transit OD Data Collection with Occlusion-Resistant Dynamic Passenger Re-Identification

HardwareDGX agent

arXiv:2504.11500v3 Announce Type: replace-cross Abstract: Transit Origin-Destination (OD) data are fundamental for optimizing public transit services, yet current collection methods, such as manual su

Tree-of-Ideas: Automated Research Ideation via Cross-Trajectory Reasoning over Scholarly Evolution

ResearchDGX agent

arXiv:2608.10740v1 Announce Type: new Abstract: Effective research ideation requires moving beyond a static understanding of prior work to trace how research problems and solutions evolve across the l

Try Grok 4.6 on tough real-world tasks!

Model ReleasesDGX agent

Try Grok 4.6 on tough real-world tasks! imo GDPVal is probably the most important benchmark, it measures the performance of models on real world tasks Big leap in performance here to top it at a great

Try Qwen-Image-3.0 on @openart_ai! 🎨👀

Model ReleasesDGX agent

Try Qwen-Image-3.0 on @openart_ai! 🎨👀 Qwen Image 3.0 is now on OpenArt ✨ The most Real Qwen image model yet. Native text across 12 languages, precise 10px type, and full interfaces like web pages, gam

TSCoNet: A Two-Stage Copula CNN-LSTM for Uncertainty-Aware Spatio-Temporal Forecasting

ResearchDGX agent

arXiv:2607.10410v2 Announce Type: replace-cross Abstract: Reliable forecasting of several interrelated environmental variables - such as regional precipitation and temperature, or other correlated geo

Twitch streamers can now opt out from training Amazon’s AI

IndustryDGX agent

Twitch users can now opt out of allowing their content to be used to train Amazon's generative AI models. Opting out means that 'your streams, VODs, clips, stream chats, and pictures and text on your

Two-stage Odd Residual Flows for Mean-Preserving Probabilistic Time Series Forecasting

ResearchDGX agent

arXiv:2608.11114v1 Announce Type: cross Abstract: Probabilistic forecasting plays an essential role in risk-sensitive decision-making, particularly in long-horizon settings. However, existing approach

Uncertainty-Aware Deep Learning for Genomics Applications: Insights from an Empirical Study

ResearchDGX agent

arXiv:2608.11054v1 Announce Type: new Abstract: Deep learning models have emerged as the standard computational tool for a wide range of applications in genomics. Yet, uncertainty quantification (UQ)

Uncertainty-Aware Ensemble Deep Randomized Neural Networks for Classification

Model ReleasesDGX agent

arXiv:2608.10007v1 Announce Type: cross Abstract: The current state-of-the-art (SOTA) deep randomized neural networks, such as deep Random Vector Functional Link (dRVFL) and ensemble deep RVFL (edRVFL

UniMod: Enhancing Multi-Modal Medical Diagnosis through Cross-Modality and Within-Modality Alignment

SafetyDGX agent

arXiv:2608.10316v1 Announce Type: new Abstract: Multi-modal learning combining medical images and clinical text is promising for disease diagnosis. However, standard multi-modal training leads to shor

UniProbe: A Learnable Token-Level Hallucination Detector for Large VLMs using Multi-Structural Internal Representations

Local AiDGX agent

arXiv:2608.10835v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) achieve impressive visual reasoning and dialogue capabilities, yet frequently hallucinate content unsupported by th

Unlocking the Power of Medical Tabular Data via Semantic-Aware Multimodal Pre-training

ResearchDGX agent

arXiv:2608.10522v1 Announce Type: cross Abstract: While vision-language models dominate medical representation learning, unstructured text lacks the dense, quantitative diagnostic phenotypes inherent

Unsupervised Detection of Groundwater Storage Anomalies in Ghana Using GRACE Satellite Data

ResearchDGX agent

arXiv:2608.10233v1 Announce Type: cross Abstract: Groundwater variability in Ghana remains poorly characterized due to limited long-term in-situ observations. This study investigates groundwater stora

UPAIR: Diagnosing Reasoning States via Uncertainty-Progress Alignment for Selective Intervention

SafetyDGX agent

arXiv:2607.17188v2 Announce Type: replace Abstract: While test-time scaling improves the problem-solving ability of large reasoning models (LRMs) through additional inference-time computation, it can

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs

Model ReleasesDGX agent

arXiv:2608.10042v1 Announce Type: cross Abstract: Tool-use LLMs are increasingly asked to act on users' behalf, but existing benchmarks usually focus on profile recall, style imitation, generic tool u

UT-ACA: Uncertainty-Triggered Adaptive Context Allocation for Long-Context Inference

Model ReleasesDGX agent

arXiv:2603.18446v2 Announce Type: replace Abstract: Long-context inference remains challenging for large language models due to attention dilution and out-of-distribution degradation. Context selectio

V-FiLLM: Verified Financial LLM Reasoning Benchmark

Model ReleasesDGX agent

arXiv:2608.11047v1 Announce Type: new Abstract: While existing benchmarks have made substantial progress in evaluating LLMs across STEM domains, financial reasoning over structured data remains compar

v0.32.10-rc0: nn: speed up prefill on double-scale nvfp4 models

Local AiDGX agent

ModelOpt checkpoints apply a float32 global scale to every projection output on top of the per-group quantization scales. Running the multiply and the cast back to the activation dtype as separate eag

Validated Synthetic Patient Generation for Small Longitudinal Cohorts: Coagulation Dynamics Across Pregnancy

ResearchDGX agent

arXiv:2604.07557v2 Announce Type: replace Abstract: Small longitudinal cohorts, common in maternal health, rare diseases, and early-phase trials, limit computational modeling because enrollment is slo

VDC-Agent: When Video Detailed Captioners Evolve Themselves via Agentic Self-Reflection

AgentsDGX agent

arXiv:2511.19436v2 Announce Type: replace-cross Abstract: Existing Video Detailed Captioning (VDC) methods predominantly rely on costly human annotations or distillation from powerful proprietary mode

VERDICT: Training-Free Step-Wise Verification of Multimodal Reasoning via Disagreement-Aware Consensus

ResearchDGX agent

arXiv:2608.10665v1 Announce Type: new Abstract: Multimodal large language models often generate reasoning chains containing subtle errors that lead to incorrect answers. Current verification approache

Very interesting new work from Anthropic. (bookmark it) They evolve mind viruses, ideas that spread through a multi-agent system by getting …

AgentsDGX agent

Very interesting new work from Anthropic. (bookmark it) They evolve mind viruses, ideas that spread through a multi-agent system by getting each host to pass them on, then measure what governs the spr

VibeLifeBench: Can Your Life Agent Be Proactive and Persistent in a Living World?

Model ReleasesDGX agent

arXiv:2608.10875v1 Announce Type: cross Abstract: Large language model (LLM) agents are increasingly deployed as personal assistants. Existing evaluations, however, mostly use short, self-contained re

VidForensics-M1: Meta-Detection Reinforcement Learning with Verifiable Temporal Grounding for AI-Generated Video Forensics

Local AiDGX agent

arXiv:2608.11201v1 Announce Type: new Abstract: Recent advances in video generation models have significantly improved the realism of synthetic videos, blurring the boundary between generated and auth

VIDS-Seg: Towards Reliable Uncertainty Quantification in Pediatric Cardiac Ultrasound Segmentation

SafetyDGX agent

arXiv:2608.10903v1 Announce Type: new Abstract: Reliable clinical deployment of machine learning requires models that know when they are likely to fail, particularly for subgroups underrepresented in

VIScore: Diagnosing Planning-Relevant Quality in Latent World Models

ResearchDGX agent

arXiv:2608.11174v1 Announce Type: new Abstract: Regulating the latent space to an isotropic Gaussian distribution provides a stable and information-maximized landscape for world model planning. Howeve

VisEditBench: Can Vision-Language Models Edit Visualization Code from Multimodal Feedback?

Model ReleasesDGX agent

arXiv:2608.10408v1 Announce Type: new Abstract: Vision-language models (VLMs) have shown strong capabilities in generating visualization code from textual or visual specifications. However, real-world

Vision-Language-Motion Maps: An Open-Vocabulary, Uncertainty-Aware, Queryable Motion Attribute for 3D Scene Maps

Model ReleasesDGX agent

arXiv:2607.16173v2 Announce Type: replace Abstract: Open-vocabulary 3D maps let robots answer language queries about what and where, but they assume a static world and cannot answer queries about how

Visual Geometry Foundation-Aware Gaussians for Single-Frame Surround-View Driving Reconstruction

Model ReleasesDGX agent

arXiv:2608.10682v1 Announce Type: new Abstract: Single-frame surround-view reconstruction faces severe geometric instability and rendering artifacts due to minimal inter-camera overlap. While existing

VoxSumm: A Multilingual Corpus of Long-Form Spoken News for Joint Summarization and Translation

Model ReleasesDGX agent

arXiv:2608.10359v1 Announce Type: cross Abstract: As information increasingly traverses linguistic boundaries, users require concise cross-lingual representations of long-form content. Nevertheless, l

Watching Synthetic Videos: Aligning Cross-modal Representations with Visual Synthesis for Zero-shot Video Captioning

SafetyDGX agent

arXiv:2608.11013v1 Announce Type: new Abstract: Text-only training is a popular paradigm in zero-shot video captioning, where the video distribution is not available to the model during training, lead

WaveInst: A Frequency-Domain Enhanced Network for Fine-Grained Thin Tree Trunk Extraction in Forest Scenes

ResearchDGX agent

arXiv:2505.01656v2 Announce Type: replace Abstract: Analyzing tree morphology, particularly trunk and branch extraction, is valuable for genetic breeding and forestry management. Existing image-based

We are in an insane run of open-weight drops. Every modality, open source is winning. This is what an open source AI summer ☀️ looks like: …

Model ReleasesDGX agent

We are in an insane run of open-weight drops. Every modality, open source is winning. This is what an open source AI summer ☀️ looks like: 🧠 LLMs & Reasoning → DeepSeek-V4-Flash-0731 (my king 👑): 304B

← Previous
1…910111213…1387
Next →