AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
84,570 results
3 Aug 2026

.@ssankar says Palantir was able to make Nvidia's Nemotron Ultra model 'better than frontier': 'I literally almost felt gaslit when, within …

Model ReleasesDGX agent

.@ssankar says Palantir was able to make Nvidia's Nemotron Ultra model 'better than frontier': 'I literally almost felt gaslit when, within 24 hours of getting Nemotron up with no post-training, this

ST-WAM: Semantic-Temporal World Action Model for Robust Manipulation under Visual Distribution Shifts

ApplicationsDGX agent

arXiv:2607.28993v1 Announce Type: cross Abstract: World Action Models (WAMs) have emerged as a promising paradigm by jointly modeling robot actions and future visual dynamics. However, their reliance

Stable Autoregressive Speech Generation with Low-Frame-Rate High-Dimensional Continuous Tokens

ResearchDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub

arXiv:2607.29363v1 Announce Type: cross Abstract: Balancing sequence length, representational capacity, and long-horizon stability is a central problem in autoregressive (AR) speech and audio generati

STAGE: STyle-controllable Action GEneration for personalized autonomous driving

SafetyDGX agent

arXiv:2607.29517v1 Announce Type: new Abstract: Driving style refers to the behavioral preferences that drivers maintain during driving, shaped by their diverse experiences, habits, and needs, and is

StaQ: a Finite Memory Approach to Discrete Action Policy Mirror Descent

SafetyDGX agent

arXiv:2506.13862v2 Announce Type: replace-cross Abstract: In Reinforcement Learning (RL), regularization with a Kullback-Leibler divergence that penalizes large deviations between successive policies

Stem: Rethinking Causal Information Flow in Sparse Attention

ResearchDGX agent

arXiv:2603.06274v2 Announce Type: replace-cross Abstract: The quadratic computational complexity of self-attention remains a fundamental bottleneck for scaling Large Language Models (LLMs) to long con

Step-Level Visual Grounding Faithfulness Predicts Out-of-Distribution Generalization in Long-Horizon Vision-Language Models

Model ReleasesDGX agent

arXiv:2603.06828v2 Announce Type: replace-cross Abstract: We uncover a behavioral law of long-horizon vision-language models: models that maintain temporally grounded beliefs generalize better. Standa

StraightDP: Geometry-Aware Differential Privacy for Rectified-Flow Transformers

Model ReleasesDGX agent

arXiv:2607.29100v1 Announce Type: cross Abstract: Differentially private (DP) training of text-conditioned generative models suffers a utility cliff at strong privacy. We revisit this problem through

Stratified Negation in RDF Rules: A Correct Approach (Extended Version)

SafetyDGX agent

arXiv:2607.28778v1 Announce Type: cross Abstract: Combining RDF rule languages, such as N3 or SHACL Rules, with default negation is challenging. Existing methods to stratify negation often fail for RD

Structured Neural Chaos: An Adaptive Surrogate Modeling Framework for Functional Uncertainty Quantification and Global Sensitivity Analysis

ResearchDGX agent

arXiv:2607.28903v1 Announce Type: cross Abstract: Variance-based global sensitivity analysis (GSA) plays a key role in uncertainty quantification by identifying the contributions of uncertain inputs t

Studying quantization trade-offs for efficient inference deployment in machine translation

HardwareDGX agent

arXiv:2607.29397v1 Announce Type: new Abstract: Deploying large language models in realistic server environments poses challenges, as the system needs to provide high-quality responses with low latenc

Sudden silence on metrics that used to be shared regularly is a bad sign. If this post below is correct, it doesn’t look great for enthusias…

AgentsDGX agent

Sudden silence on metrics that used to be shared regularly is a bad sign. If this post below is correct, it doesn’t look great for enthusiastic projections around Anthropic’s Q3. (Author below leaves

SULAND v2: A Refined RGB Dataset and Deep Learning Object Detection Benchmark for UAV/UGV-Based SUrface LANDmine Detection Under Domain Shift

Model ReleasesDGX agent

arXiv:2607.28996v1 Announce Type: new Abstract: RGB imagery offers a practical, low-cost option for Unmanned Aerial/Ground Vehicle (UAV/UGV) survey support in surface-landmine detection, but object de

Sycophancy Undermines Epistemic Vigilance in Cooperative Vision-Language Tasks

ResearchDGX agent

arXiv:2607.29585v1 Announce Type: new Abstract: To maintain common ground in cooperative conversation, humans iteratively update their beliefs as conversation participants share new information; parti

Symplectic Representation of Legendre Dynamics

ResearchDGX agent

arXiv:2512.19409v2 Announce Type: replace Abstract: Modern learning systems act on internal representations of data, yet how these representations encode underlying physical or statistical structure i

TacPrint: A Wearable Fingertip Tactile Sensor for Human-to-Robot Contact Reproduction

Local AiDGX agent

arXiv:2607.29231v1 Announce Type: new Abstract: Human-centric data collection is emerging as a significant paradigm for robot skill acquisition, but seamlessly integrating low-cost, scalable tactile s

TAGTorch: A PyTorch Library for Geometry, Topology, and Symmetry-Aware Machine Learning

ResearchDGX agent

arXiv:2607.28755v1 Announce Type: new Abstract: Over the last decade, neural networks have been applied to an increasingly diverse range of applications, including data with rich geometric, topologica

TAPR: Enhancing LLM Performance with a Task-Aware Prompt Rewriter

Model ReleasesDGX agent

arXiv:2607.28657v1 Announce Type: new Abstract: Large Language Models (LLMs) often require carefully crafted prompts to unlock their full potential, which can be a barrier for non-expert users. This w

TAVI-TEC: An AI-Based Tool for Procedural Planning of Transcatheter Aortic Valve Implantation

ResearchDGX agent

arXiv:2607.29243v1 Announce Type: cross Abstract: Computed tomography angiography (CTA) is crucial for preprocedural TAVI planning, providing the anatomical information required for prosthesis sizing

Teaching Video Generators to Remember: Eliciting Dynamic Memory for Out-of-Sight State Evolution

Model ReleasesDGX agent

arXiv:2605.25333v2 Announce Type: replace Abstract: Video world models should maintain evolving states when evidence is unobserved, yet current generators often freeze hidden states upon interruption.

Technological Advances in Detecting and Managing Cognitive Impairment in Older Adults: Trends, Challenges, and Future Directions

ResearchDGX agent

arXiv:2607.28687v1 Announce Type: cross Abstract: As populations age, cognitive decline from mild cognitive impairment (MCI) to dementia is a defining health challenge of the coming decades, yet routi

TELLER: Dual-Path Iterative Preference Optimization for Table Entity Linking

SafetyDGX agent

arXiv:2607.28680v1 Announce Type: new Abstract: Entity linking in tables matches short and ambiguous cell mentions to their corresponding knowledge-base entities. Existing approaches typically rely on

Temporal Policy: History-Initialized Action Generation for Robotic Learning from Demonstration

SafetyDGX agent

arXiv:2607.29482v1 Announce Type: new Abstract: By relying on independent couplings from uninformative Gaussian priors, standard diffusion and flow matching models are forced to learn complex, high-co

Tensor Data Scattering and the Impossibility of Slicing Theorem

ResearchDGX agent

arXiv:2012.01982v3 Announce Type: replace Abstract: This paper proposes a standard way to represent sparse tensors. A broad theoretical framework for tensor data scattering methods used in various dee

TerraNova: A Foundation Model for the Anthropocene

SafetyDGX agent

arXiv:2607.29527v1 Announce Type: cross Abstract: A defining problem of the Anthropocene is to model the physical Earth and human societies as one coupled system, yet no learned representation spans t

TextCloak: Thwarting Unauthorized LLM Exploitation via RL-Driven Unlearnable Text

SafetyDGX agent

arXiv:2607.28862v1 Announce Type: cross Abstract: The rapid development of Large Language Models (LLMs) has led to significant advances across a wide range of language tasks, while simultaneously rais

TFGformer: Multivariate Time Series Forecasting via Time-Frequency Graph Learning and Covariate Fusion

ResearchDGX agent

arXiv:2607.29459v1 Announce Type: cross Abstract: Large-scale multivariate time series from heterogeneous IoT sensors demand accurate long-term forecasting for resource scheduling and predictive maint

Thanks for the recognition. We'll keep building! 🚀

Model ReleasesDGX agent

Thanks for the recognition. We'll keep building! 🚀 Big news: Qwen3.8-Max by @Alibaba_Qwen just landed at #4 on the Frontend Code Arena leaderboard with a score of 1,668! With 1,668 points, Qwen3.8-Max

The Asymmetric Effects of Knowledge Distillation on Bias in Small Language Models

Model ReleasesDGX agent

arXiv:2607.28639v1 Announce Type: cross Abstract: We show that knowledge distillation in small instruction-tuned language models has asymmetric effects on bias. On unambiguous tasks (BBQ-disambig), re

The Capability Convergence Hypothesis: Capability from Access Structure, Not Scale

ResearchDGX agent

arXiv:2607.14144v2 Announce Type: replace Abstract: The Platonic Representation Hypothesis (PRH) holds that as models scale, representations of heterogeneous networks converge toward a shared model of

The Checking Problem: What must be true before AI ships in a regulated firm

ApplicationsDGX agent

arXiv:2607.28666v1 Announce Type: new Abstract: Enterprise AI programmes stall at a rate that is widely quoted and poorly explained. This paper measures the mechanism. Six document-heavy workflows of

The Chinese labs everyone lumps together are making four pretty different bets. I work at one of them.

Model ReleasesDGX agent

Every time a model drops from a Chinese lab the thread fills with people who already know who made it, and the guess is usually Alibaba. There was a thread here recently asking what separates the open

The desktop app UI is getting a massive upgrade. What's next on the roadmap?

Local AiDGX agent

I've been following the recent pull requests and saw that the desktop app is being transformed from a chat-only interface into a full management tool with a tabbed settings UI, a model manager, and a

The Download: reward hacking explained, and suspected Iranian cyberattacks

ResearchDGX agent

This is today’s edition of The Download, our weekday newsletter that provides a daily dose of what’s going on in the world of technology. Here’s why AI agents lie and cheat to reach their goals When t

The Formalism Trap: Are LLM-as-a-Judge Evaluators Blinded by Consensus Mimicry under Social Load?

AgentsDGX agent

arXiv:2607.28641v1 Announce Type: cross Abstract: We introduce the extit{Agentic Formalism Trap} and the Evaluative Dissonance Index (D_E), quantifying how LLM-as-a-Judge systems conflate structural p

The Greedy Advantage in Finite-Horizon Bandits

SafetyDGX agent

arXiv:2607.29375v1 Announce Type: cross Abstract: Organizations increasingly rely on sequential experimentation to improve decision-making. While the multi-armed bandit literature has developed algori

The Grokked Illusion: True Equilibrium Mitigates Catastrophic Forgetting

Model ReleasesDGX agent

arXiv:2607.29503v1 Announce Type: new Abstract: While neural networks are typically evaluated by their training and test performance, these metrics do not reveal how robust a learned representation is

The Inference Engineering Masterclass: 10x faster models, quantization, speculative decoding, Rubin, & self-optimizing AI https://www.latent…

HardwareDGX agent

The Inference Engineering Masterclass: 10x faster models, quantization, speculative decoding, Rubin, & self-optimizing AI https://www.latent.space/p/inference-eng @Baseten @philipkiely and @waterloo_i

The K-Space Signature: Frequency-Domain Representation Learning for Medical Deepfake Detection

SafetyDGX agent

arXiv:2607.29541v1 Announce Type: new Abstract: In medical imaging, generative models are increasingly deployed to synthesize realistic data and augment limited datasets. Unfortunately, while benefici

The math still ain’t mathing.

HardwareDGX agent

The math still ain’t mathing. Recently we estimated global (ex China) AI revenues of around 200bn annualised, based on four different sources. Just come across a fifth, based on Nvidia inference sales

The Morphological Core of Dungan: A Two-Dialect Finite-State Model and a Multi-Genre Evaluation

Model ReleasesDGX agent

arXiv:2607.28766v1 Announce Type: new Abstract: Dungan, a Sinitic language of Central Asia written in a Cyrillic-based script, is described in detail in the grammatical literature, yet the quantitativ

The New Monday Morning Report: How Generative AI can deliver the insights your executives need.

IndustryDGX agent

The article discusses the launch of a revamped Monday Morning Report powered by generative AI, designed to automatically aggregate, analyze, and present key business metrics to executives. By leveragi

The Parts Are Greater Than the Sum: Automated Task Sequencing for Efficient Training of Multi-Policy LLMs

Model ReleasesDGX agent

arXiv:2607.29601v1 Announce Type: new Abstract: Parameter-Efficient Fine-Tuning (PEFT) commonly adapts large language models using a single shared Low-Rank Adapter (LoRA). This shared optimization spa

The persuasive power of large language models does not depend on their perceived national origin

ResearchDGX agent

arXiv:2607.29334v1 Announce Type: cross Abstract: Conversational AI developed by geopolitical rivals reaches citizens worldwide, raising concerns that it could sway public opinion or be rejected as fo

The result is a faster, more natural conversation with ChatGPT Voice from the moment a session starts. How we built it: https://openai.com/i…

Model ReleasesDGX agent

OpenAI has redesigned the ChatGPT Voice stack—from client to model—to enable continuous audio streaming, allowing GPT‑Live to listen while speaking without interruption. The new architecture supports

The results span sphere packing, coding theory, group theory, quantum complexity, lattice cryptography, extremal combinatorics, and more. Am…

Model ReleasesDGX agent

The results span sphere packing, coding theory, group theory, quantum complexity, lattice cryptography, extremal combinatorics, and more. Among them: establishing the existence of non-sofic groups and

The Theoretical Foundation of Socratic Tests: Dynamic, Multimodal, Conversational Examinations

SafetyDGX agent

arXiv:2607.29624v1 Announce Type: cross Abstract: Traditional static assessments rely on a subtractive, deficit-based grading model that often penalizes ambition and obscures diagnostic feedback. Conv

There should be a public investigation into the AI hacking incidents by OpenAI and Anthropic. We deserve to know whether these labs are genu…

SafetyDGX agent

There should be a public investigation into the AI hacking incidents by OpenAI and Anthropic. We deserve to know whether these labs are genuinely world-class security organizations facing a novel thre

ThinkReset: Learnable Intermediate Interface Construction for Bounded-Context Long-Horizon Reasoning

ResearchDGX agent

arXiv:2607.28642v1 Announce Type: new Abstract: Long chain-of-thought reasoning improves performance on complex problems, but it also introduces redundancy accumulation, context overflow, and error an

This is a wild result. Locus, the automated research system from @intology, post-trained Qwen3 base models that beat the official human-tune…

ApplicationsDGX agent

This is a wild result. Locus, the automated research system from @intology, post-trained Qwen3 base models that beat the official human-tuned Qwen3 1.7B Instruct release. SoTA on PostTrainBench! The m

this is the real reason people from OpenAI etc are desperate to shut me up. Astra (which didn’t even have a control group and is maybe not t…

SafetyDGX agent

this is the real reason people from OpenAI etc are desperate to shut me up. Astra (which didn’t even have a control group and is maybe not that much better than Fable and certainly not ASI) was perhap

Tipping Point Forecasting in Non-Stationary Dynamics on Function Spaces

TutorialsDGX agent

arXiv:2308.08794v4 Announce Type: replace Abstract: Tipping points are abrupt, drastic, and often irreversible changes in the evolution of non-stationary and chaotic dynamical systems. For instance, i

To Add Is Machine, To Delete Is Human: Measuring and Mitigating Deletion Avoidance in LLM Code Editing

Model ReleasesDGX agent

arXiv:2607.28887v1 Announce Type: cross Abstract: Large language models increasingly write and repair production code, yet evidence is mounting that their test-passing patches leave codebases harder t

To Facilitate or not to Facilitate: Human and LLM Facilitator Tendencies in Online Discussions

TutorialsDGX agent

arXiv:2607.28643v1 Announce Type: cross Abstract: Automating facilitation in online discussions is a long-standing social concern given the increasing time we spend on online spaces and the failure of

Token-Level Diagnosis of Sycophancy in LLMs with Attribution-Guided Steering

ResearchDGX agent

arXiv:2607.28906v1 Announce Type: new Abstract: Sycophancy refers to the tendency for large language models (LLMs) to match user beliefs at the cost of factual correctness, thereby undermining model r

Tokenizer-Agnostic Engram Module

Model ReleasesDGX agent

arXiv:2607.29065v1 Announce Type: new Abstract: Deepseek's Engram, a conditional memory module, was introduced to trade-off storage versus reasoning in large language models. However, the module relie

Tokenizer Transplantation: Mitigating Autoregressive Collapse in Edge-Efficient Bengali ASR

ResearchDGX agent

arXiv:2607.09598v2 Announce Type: replace Abstract: Lightweight speech recognition models are critical for edge deployment, yet highly optimized architectures like Moonshine often fail on morphologica

TokenSwap: Benchmarking and Reducing the Modality Gap in Multimodal LLMs

ResearchDGX agent

arXiv:2607.28640v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) should generate consistent responses given semantically equivalent inputs across modalities. However, we observ

TokTier: Exact Stateful Tokenization for Agentic LLM Serving

HardwareDGX agent

arXiv:2607.29678v1 Announce Type: new Abstract: LLM serving systems cache prompt KV state, yet most front ends still re-tokenize the full request text on every call. The cost lands on coding agents, w

TOOD: Task-Aware Out-of-Distribution Score Calibration for Continual Learners

TutorialsDGX agent

arXiv:2607.29592v1 Announce Type: new Abstract: The primary challenge of continual learning (CL) systems is to learn new tasks while remaining performant on previously learned tasks. A similarly impor

← Previous
1…135136137138139…1410
Next →