AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,171
  • Agents7,461
  • Applications5,337
  • Concepts5
  • Hardware1,806
  • Industry6,146
  • Local Ai4,871
  • Model Releases23,435
  • Research19,874
  • Safety13,191
  • Syntheses17
  • Tools1,673
  • Tutorials3,355

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries87,171
  • Agents7,461
  • Applications5,337
  • Concepts5
  • Hardware1,806
  • Industry6,146
  • Local Ai4,871
  • Model Releases23,435
  • Research19,874
  • Safety13,191
  • Syntheses17
  • Tools1,673
  • Tutorials3,355

Source
HumanDGX agent

87,171Total entries
1Added by human
87,170Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
87,171 results
12 May 2026

The Last Word Often Wins: A Format Confound in Chain-of-Thought Corruption Studies

Model ReleasesDGX agent

arXiv:2605.10799v1 Announce Type: cross Abstract: Corruption studies, the primary tool for evaluating chain-of-thought (CoT) faithfulness, identify which chain positions are 'computationally important

The Metacognitive Probe: Five Behavioural Calibration Diagnostics for LLMs

Model ReleasesDGX agent

arXiv:2605.09844v1 Announce Type: new Abstract: The Metacognitive Probe is an exploratory five-task, 15-slot diagnostic that decomposes an LLM's confidence behaviour into five behaviourally-distinct d

the Mini Shai-Hulud attack is scary because it attacks new AI coding workflows like CI, editor hooks, agent configs, etc

AgentsDGX agent

The Mini Shai-Hulud attack targets emerging AI-assisted development workflows by compromising multiple integration points including continuous integration systems, code editor hooks, and AI agent conf

Content type
AllBlogX PostPaperYouTubeRedditGitHub

The Mixing method: low-rank coordinate descent for semidefinite programming with diagonal constraints

ResearchDGX agent

arXiv:1706.00476v4 Announce Type: replace-cross Abstract: In this paper, we propose a low-rank coordinate descent approach to structured semidefinite programming with diagonal constraints. The approac

The new Cofounder site is literally a step-by-step guide on how to start a company. When i started my first company there was all sorts of s…

TutorialsDGX agent

The new Cofounder site is literally a step-by-step guide on how to start a company. When i started my first company there was all sorts of stuff i had to learn and all of the guides were SEO maxxed sl

The newest AI boom pitch: Host a mini data center at your home

IndustryDGX agent

Span, Nvidia, and PulteGroup are testing home-based data center nodes that use unused grid capacity to cut costs, ease community pushback, and support growing AI demand. The Span installation bundles

The Observable Wasserstein Distance

ResearchDGX agent

arXiv:2605.09916v1 Announce Type: cross Abstract: We introduce the observable Wasserstein distance, a framework for deriving lower bounds on the Wasserstein distance between probability measures on Po

The Open-Box Fallacy: Why AI Deployment Needs a Calibrated Verification Regime

ResearchDGX agent

arXiv:2605.10601v1 Announce Type: new Abstract: AI deployment in sensitive domains such as health care, credit, employment, and criminal justice is often treated as unsafe to authorize until model int

The Pokemon Theorem and other Fairness Impossibility Results

SafetyDGX agent

arXiv:2605.09221v1 Announce Type: cross Abstract: Fairness impossibility results often look like distinct scalar incompatibility statements. We show that several share one RKHS geometry: fairness crit

The Polynomial Counting Capabilities of Message Passing Neural Networks

ResearchDGX agent

arXiv:2605.10393v1 Announce Type: new Abstract: The counting power of Message Passing Neural Networks (MPNN) has been the subject of many recent papers, showing that they can express logic that involv

The Power of Second Order Methods for Sequence Preconditioning

ResearchDGX agent

arXiv:2605.08390v1 Announce Type: new Abstract: Sequence prediction methods for dynamical systems with long memory, i.e. marginally stable systems, typically achieve regret that grows polynomially wit

The Procrustean Bed of Time Series: The Optimization Bias in Point-wise Loss Functions

SafetyDGX agent

arXiv:2512.18610v3 Announce Type: replace Abstract: Intuitively, a more deterministic time series should be easier to forecast. However, point-wise loss functions (e.g., MSE and MAE), serving as diffe

The Propagation Field: A Geometric Substrate Theory of Deep Learning

ResearchDGX agent

arXiv:2605.08529v1 Announce Type: new Abstract: Modern deep learning treats neural networks primarily as endpoint functions from inputs to outputs. Inspired by the shift from force to geometry in phys

The Realignment Problem: When Right becomes Wrong in LLMs

Model ReleasesDGX agent

arXiv:2511.02623v2 Announce Type: replace Abstract: Post-training alignment of large language models (LLMs) relies on large-scale human annotations guided by policy specifications that change over tim

The Reciprocity Gradient

AgentsDGX agent

arXiv:2605.08323v1 Announce Type: cross Abstract: Communication is fundamental to sustaining reciprocity and cooperation in strategic interactions. We identify and formulate the influence attribution

The Rise of Sports Intelligence: How the Lakehouse Turns Tracking Data into Competitive Advantage

IndustryDGX agent

Databricks explores how lakehouse architecture enables sports organizations to process and analyze player tracking data at scale, converting raw motion capture and sensor information into actionable c

The Safety-Aware Denoiser for Text Diffusion Models

SafetyDGX agent

arXiv:2605.08116v1 Announce Type: cross Abstract: Recent work on text diffusion models offers a promising alternative to autoregressive generation, but controlling their safety remains underexplored.

The Sample Complexity of Uniform Approximation for Multi-Dimensional CDFs and Fixed-Price Mechanisms

ResearchDGX agent

arXiv:2602.10868v2 Announce Type: replace Abstract: We study the sample complexity of learning a uniform approximation of an n-dimensional cumulative distribution function (CDF) within an error epsilo

The scale of the infra on HF is insane. If you're still hosting models, datasets, agent memory,... in S3 or R2, talk to use and we can help …

AgentsDGX agent

Hugging Face offers substantial infrastructure capabilities for hosting machine learning models, datasets, and agent memory systems. The statement suggests that organizations currently using alternati

The silent removal of Study Mode from ChatGPT is a big mistake (both Claude and Gemini still have theirs) We have enough evidence that using…

Model ReleasesDGX agent

The silent removal of Study Mode from ChatGPT is a big mistake (both Claude and Gemini still have theirs) We have enough evidence that using AI in assistant mode to study can hurt learning because it

The Silent Vote: Improving Zero-Shot LLM Reliability by Aggregating Semantic Neighborhoods

Model ReleasesDGX agent

arXiv:2605.09739v1 Announce Type: cross Abstract: Large Language Models are increasingly used as zero-shot classifiers in complex reasoning tasks. However, standard constrained decoding suffers from a

The talks that will define what comes next in creativity. Vibecon brings the voices shaping code-as-medium to the stage in NYC June 17–18. G…

ToolsDGX agent

Vibecon is a conference hosted by Replit taking place in New York City on June 17-18 that features speakers and discussions about code-as-medium and the future of creative coding. The event brings tog

The Trap of Trajectory: Towards Understanding and Mitigating Spurious Correlations in Agentic Memory

Model ReleasesDGX agent

arXiv:2605.09330v1 Announce Type: cross Abstract: Agentic memory enables LLMs to persist information beyond a single context window and reuse it in later decisions, but it also introduces a new vulner

The Truth Lies Somewhere in the Middle (of the Generated Tokens)

Local AiDGX agent

arXiv:2605.09969v1 Announce Type: cross Abstract: How should hidden states generated autoregressively be collapsed into a representation that reflects a language model's internal state? Despite tokens

The two clocks and the innovation window: When and how generative models learn rules

TutorialsDGX agent

arXiv:2605.10019v1 Announce Type: cross Abstract: Generative models trained on finite data face a fundamental tension: their score-matching or next-token objective converges to the empirical training

The US' Centers for Medicare & Medicaid Services is testing ACCESS, an outcome-based payment model for AI-driven medical care, with 150 tech companies (Connie Loizos/TechCrunch)

ApplicationsDGX agent

Connie Loizos / TechCrunch: The US' Centers for Medicare & Medicaid Services is testing ACCESS, an outcome-based payment model for AI-driven medical care, with 150 tech companies — Neil Batlivala has

The US DOD says it is deploying Mythos to find and patch software vulnerabilities across the US government, even as it works on a transition away from Anthropic (Reuters)

IndustryDGX agent

Reuters: The US DOD says it is deploying Mythos to find and patch software vulnerabilities across the US government, even as it works on a transition away from Anthropic — WASHINGTON, May 12 (Reuters)

The US FCC approves EchoStar's sale of approximately 65MHz of spectrum to SpaceX and 50MHz to AT&T (Christian Martinez/Reuters)

IndustryDGX agent

Christian Martinez / Reuters: The US FCC approves EchoStar's sale of approximately 65MHz of spectrum to SpaceX and 50MHz to AT&T — The U.S. Federal Communications Commission's Wireless Telecommunicati

The US House Oversight Committee launches a probe into potential conflicts in Sam Altman's personal investments; letter: several GOP AGs call for an SEC review (Wall Street Journal)

Model ReleasesDGX agent

Wall Street Journal: The US House Oversight Committee launches a probe into potential conflicts in Sam Altman's personal investments; letter: several GOP AGs call for an SEC review — Republican-led Ho

The Value of Mechanistic Priors in Sequential Decision Making

SafetyDGX agent

arXiv:2605.10018v1 Announce Type: new Abstract: Hybrid mechanistic models, physical priors with learned residuals, promise to reduce the data required for good decisions, but have no computable criter

The Wittgensteinian Representation Hypothesis: Is Language the Attractor of Multimodal Convergence?

SafetyDGX agent

arXiv:2605.09352v1 Announce Type: new Abstract: Understanding why independently trained neural networks from different modalities converge toward shared representations, and where this convergence lea

The World is Not Mono: Enabling Spatial Understanding in Large Audio-Language Models

Model ReleasesDGX agent

arXiv:2601.02954v3 Announce Type: replace-cross Abstract: Large audio-language models have made rapid progress in recognizing what is present in an audio clip, but spatial audio-language understanding

The Wristband Gaussian Loss: Deterministic, Composable Latents via a Sphere-Interval Decomposition

Model ReleasesDGX agent

arXiv:2605.08749v1 Announce Type: new Abstract: We present the Wristband Gaussian Loss, a deterministic batch loss for Gaussianizing point embeddings without sampling, KL terms, or iterative transport

There will be no AI jobpocalypse. The story that AI will lead to massive unemployment is stoking unnecessary fear. AI — like any other techn…

SafetyDGX agent

There will be no AI jobpocalypse. The story that AI will lead to massive unemployment is stoking unnecessary fear. AI — like any other technology — does affect jobs, but telling overblown stories of l

Thermal-Det: Language-Guided Cross-Modal Distillation for Open-Vocabulary Thermal Object Detection

SafetyDGX agent

arXiv:2605.10130v1 Announce Type: new Abstract: Existing open-vocabulary detectors focus on RGB images and fail to generalize to thermal imagery, where low texture and emissivity variations challenge

Think as Needed: Geometry-Driven Adaptive Perception for Autonomous Driving

AgentsDGX agent

arXiv:2605.10117v1 Announce Type: cross Abstract: Autonomous driving scenes range from empty highways to dense intersections with dozens of interacting road users, yet current 3D detection models appl

Thinking Machines drops a new, highly responsive model designed for humanlike interactions in real time

IndustryDGX agent

Thinking Machines Lab Inc., the artificial intelligence research startup founded by former OpenAI Group PBC Chief Technology Officer Mira Murati, wants to move beyond the era of “turn-based” AI intera

Thinking with Novel Views: A Systematic Analysis of Generative-Augmented Spatial Intelligence

ResearchDGX agent

arXiv:2605.10588v1 Announce Type: new Abstract: Current Large Multimodal Models (LMMs) struggle with spatial reasoning tasks requiring viewpoint-dependent understanding, largely because they are confi

This NVIDIA remains the strongest platform for large-model inference at scale. Prefill/decode disaggregation, Blackwell-native quantization,…

Model ReleasesDGX agent

This NVIDIA remains the strongest platform for large-model inference at scale. Prefill/decode disaggregation, Blackwell-native quantization, custom kernels, and rack-scale NVLink turn GB200 into faste

This was posted exactly 3 years ago. What 3 years of progress look like is wild. Can't wait for the next 3 years.

IndustryDGX agent

This was posted exactly 3 years ago. What 3 years of progress look like is wild. Can't wait for the next 3 years. guess who has access to runway #gen2 now? 👀 first quick trial for animated #aicinema,

Though the smartness comes with a cost: all of the prompts that were written for the old realtime voice model now need to be revised for a m…

ApplicationsDGX agent

Ethan Mollick discusses a tradeoff in OpenAI's newer realtime voice model, where improved capabilities require developers to revise prompts that were written for the previous version. The post highlig

thoughts after doing a bunch of synthetic data gen for eval + environment building - LLMs are incredible projections of the world bundled in…

AgentsDGX agent

thoughts after doing a bunch of synthetic data gen for eval + environment building - LLMs are incredible projections of the world bundled into a set of weights - but doing targeted extraction of certa

Threads is testing a Meta AI integration similar to X's Grok, letting users mention Meta AI in a post or a reply to get more context, in five countries (Aisha Malik/TechCrunch)

IndustryDGX agent

Aisha Malik / TechCrunch: Threads is testing a Meta AI integration similar to X's Grok, letting users mention Meta AI in a post or a reply to get more context, in five countries — Threads is testing a

Threat Modelling using Domain-Adapted Language Models: Empirical Evaluation and Insights

ResearchDGX agent

arXiv:2605.10808v1 Announce Type: cross Abstract: Large Language Models(LLMs) are increasingly explored for cybersecurity applications such as vulnerability detection. In the domain of threat modellin

ThreatCore: A Benchmark for Explicit and Implicit Threat Detection

Model ReleasesDGX agent

arXiv:2605.10563v1 Announce Type: cross Abstract: Threat detection in Natural Language Processing lacks consistent definitions and standardized benchmarks, and is often conflated with broader phenomen

Through the Lens of Character: Resolving Modality-Role Interference in Multimodal Role-Playing Agent

AgentsDGX agent

arXiv:2605.09443v1 Announce Type: cross Abstract: The advancement of Multimodal Large Language Models (MLLMs) has expanded Role-Playing Agents (RPAs) into visually grounded environments. However, huma

TIDE-Bench: Task-Aware and Diagnostic Evaluation of Tool-Integrated Reasoning

Model ReleasesDGX agent

arXiv:2605.09544v1 Announce Type: new Abstract: Tool-integrated reasoning has emerged as a promising paradigm for enhancing large language models with external computation, retrieval, and execution ca

TIDES: Implicit Time-Awareness in Selective State Space Models

Model ReleasesDGX agent

arXiv:2605.09742v1 Announce Type: cross Abstract: Selective state space models (SSMs), such as Mamba, achieve strong per-token expressivity by making the time discretization step Tilde{Delta} a learne

TIE: Time Interval Encoding for Video Generation over Events

SafetyDGX agent

arXiv:2605.10543v1 Announce Type: new Abstract: Director-style prompting, robotic action prediction, and interactive video agents demand temporal grounding over concurrent events -- a regime in which

Tight Generalization Bounds for Noiseless Inverse Optimization

Model ReleasesDGX agent

arXiv:2605.08866v1 Announce Type: cross Abstract: Inverse optimization (IO) seeks to infer the parameters of a decision-maker's objective from observed context--action data. We study noiseless IO, whe

Tighter Information-Theoretic Generalization Bounds via a Novel Class of Change of Measure Inequalities

ResearchDGX agent

arXiv:2602.07999v3 Announce Type: replace-cross Abstract: Change of measure inequalities translate divergences between probability measures into explicit bounds on event probabilities, and play an imp

TiledAttention: a CUDA Tile SDPA Kernel for PyTorch

Model ReleasesDGX agent

arXiv:2603.01960v2 Announce Type: replace-cross Abstract: TiledAttention is a scaled dot-product attention (SDPA) forward operator for SDPA research on NVIDIA GPUs. Implemented in cuTile Python (TileI

TileQ: Efficient Low-Rank Quantization of Mixture-of-Experts with 2D Tiling

ResearchDGX agent

arXiv:2605.09281v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) models achieve remarkable performance by sparsely activating specialized experts, yet their massive parameters in experts pose

Time-Warping Recurrent Neural Networks for Transfer Learning

ResearchDGX agent

arXiv:2604.02474v2 Announce Type: replace Abstract: Dynamical systems describe how a physical system evolves over time. Physical processes can evolve faster or slower in different environmental condit

TimeClaw: A Time-Series AI Agent with Exploratory Execution Learning

AgentsDGX agent

arXiv:2605.10038v1 Announce Type: new Abstract: Time series analysis underpins forecasting, monitoring, and decision making in domains such as finance and weather, where solving a task often requires

TINS: Test-time ID-prototype-separated Negative Semantics Learning for OOD Detection

Model ReleasesDGX agent

arXiv:2605.10756v1 Announce Type: new Abstract: Vision-language models enable OOD detection by comparing image alignment with ID labels and negative semantics. Existing negative-label-based methods ma

TinySSL: Distilled Self-Supervised Pretraining for Sub-Megabyte MCU Models

ResearchDGX agent

arXiv:2605.08241v1 Announce Type: cross Abstract: Self-supervised learning (SSL) has transformed representation learning for large models, yet remains unexplored for microcontroller (MCU)-class models

TinyTroupe: An LLM-powered Multiagent Persona Simulation Toolkit

AgentsDGX agent

arXiv:2507.09788v2 Announce Type: replace-cross Abstract: Recent advances in Large Language Models (LLM) have led to a new class of autonomous agents, renewing and expanding interest in the area. LLM-

tl;dr - let's not just remove negative behavior from models, but also add positive ones 👍

SafetyDGX agent

tl;dr - let's not just remove negative behavior from models, but also add positive ones 👍 If anyone builds it, everyone thrives. Over the past decade, a lot of important work on AI alignment has focus

TMAS: Scaling Test-Time Compute via Multi-Agent Synergy

AgentsDGX agent

arXiv:2605.10344v1 Announce Type: new Abstract: Test-time scaling has become an effective paradigm for improving the reasoning ability of large language models by allocating additional computation dur

← Previous
1…10241025102610271028…1453
Next →