AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,223
  • Agents7,699
  • Applications5,506
  • Concepts5
  • Hardware1,889
  • Industry6,186
  • Local Ai5,045
  • Model Releases24,499
  • Research20,615
  • Safety13,633
  • Syntheses17
  • Tools1,677
  • Tutorials3,452

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries90,223
  • Agents7,699
  • Applications5,506
  • Concepts5
  • Hardware1,889
  • Industry6,186
  • Local Ai5,045
  • Model Releases24,499
  • Research20,615
  • Safety13,633
  • Syntheses17
  • Tools1,677
  • Tutorials3,452

Source
HumanDGX agent

90,223Total entries
1Added by human
90,222Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
90,223 results
28 May 2026

The Grammar of Transformers: A Systematic Review of Interpretability Research on Syntactic Knowledge in Language Models

ResearchDGX agent

arXiv:2601.19926v2 Announce Type: replace-cross Abstract: We present a systematic review of 337 articles evaluating the syntactic abilities of Transformer-based language models (TLMs), reporting on ov

The Harder Text Embedding Benchmark (HTEB): Beyond One-dimensional Static Robustness

Model ReleasesDGX agent

arXiv:2605.28190v1 Announce Type: new Abstract: Embedding benchmarks like MTEB report a single score per model, implicitly treating robustness as a static, scalar property. We argue that embedding rob

The HF science team just made async RL weight sync ~100x cheaper on bandwidth, and you don't need a shared cluster anymore. The problem: eve…

HardwareDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub

The HF science team just made async RL weight sync ~100x cheaper on bandwidth, and you don't need a shared cluster anymore. The problem: every RL step, the trainer typically has to sync fresh weights

The Illusion of Opting in AI-Mediated Consequential Decisions

SafetyDGX agent

arXiv:2605.28210v1 Announce Type: new Abstract: Drawing on Ullmann-Margalit's concept of opting (transformative, irrevocable, and shadowed by foreclosed alternatives), we show that current AI systems

The Importance of Being Statistically Earnest: A Critical Re-evaluation of GSM-Symbolic

Model ReleasesDGX agent

arXiv:2605.28700v1 Announce Type: new Abstract: The GSM-Symbolic benchmark (Mirzadeh et al., 2025) reported consistent performance drops across 25 Large Language Models (LLMs) when tested on template-

“The mechanism is always the same in every story I've been covering. The demo works in a controlled environment with clean inputs. The deplo…

TutorialsDGX agent

“The mechanism is always the same in every story I've been covering. The demo works in a controlled environment with clean inputs. The deployment fails because real kitchens, real intersections, and r

The Missing Piece in Pre-trained Model Evaluation: Reward-Guided Decoding Unlocks Task-Oriented Behavior Without Parameter Updates

Model ReleasesDGX agent

arXiv:2605.28020v1 Announce Type: new Abstract: With the rapid progress of large language models (LLMs), reliably evaluating the capabilities of pre-trained LLMs has become increasingly important. The

The Name’s Gaming … Cloud Gaming: ‘007 First Light’ Launches on GeForce NOW

Model ReleasesDGX agent

License to stream, shaken and stirred. GeForce NOW is dialing up the espionage with the launch of 007 First Light, letting members slip into James Bond’s reimagined origin story from almost any device

The Obfuscation Atlas: Mapping Where Honesty Emerges in RLVR with Deception Probes

SafetyDGX agent

arXiv:2602.15515v2 Announce Type: replace-cross Abstract: Training against white-box deception detectors has been proposed as a way to make AI systems honest. However, such training risks models learn

The Optimal Sample Complexity of Linear Contracts

AgentsDGX agent

arXiv:2601.01496v2 Announce Type: replace-cross Abstract: In this paper, we settle the problem of learning optimal linear contracts from data in the offline setting, where agent types are drawn from a

The Point, the Vision and the Text: Does Point Cloud Boost Spatial Reasoning of Large Language Models? A Bias-Controlled Study

Model ReleasesDGX agent

arXiv:2504.04540v2 Announce Type: replace-cross Abstract: 3D Large Language Models (LLMs) leveraging spatial information in point clouds for 3D spatial reasoning attract great attention. Despite some

The price to rent an Nvidia H200 just collapsed from 7/hr to 4/hr in three weeks. A -40% drop in the cost of the single most strategic ass…

HardwareDGX agent

The price to rent an Nvidia H200 just collapsed from 7/hr to 4/hr in three weeks. A -40% drop in the cost of the single most strategic asset in tech. When the underlying commodity that powers your ent

The Principles of Diffusion Models

TutorialsDGX agent

arXiv:2510.21890v2 Announce Type: replace-cross Abstract: This book presents the core principles that have guided the development of diffusion models, tracing their origins and showing how diverse for

The rarest object type in the universe isn't black holes. It's us. Conscious matter. The flame of life. We have a duty to expand it in scope…

IndustryDGX agent

This post expresses Elon Musk's philosophical perspective that conscious matter is exceptionally rare in the universe compared to other phenomena like black holes, and argues for a moral imperative to

The @runwayml community meetup in Tokyo was incredible! I met so many amazing people and heard so many inspiring stories. ありがとう東京、また戻ってきます! …

IndustryDGX agent

Cristobal Valenzuela, co-founder of Runway ML, attended a community meetup event in Tokyo where he connected with members of the Runway ML user community and heard personal stories about how users are

The Script is All You Need: An Agentic Framework for Long-Horizon Dialogue-to-Cinematic Video Generation

Model ReleasesDGX agent

arXiv:2601.17737v3 Announce Type: replace-cross Abstract: Recent advances in video generation have produced models capable of synthesizing stunning visual content from simple text prompts. However, th

The Shape of Overthinking: Backtracking Bursts in Long Reasoning Traces

Local AiDGX agent

arXiv:2605.27965v1 Announce Type: new Abstract: Reasoning models often generate long traces in which useful self-correction and unproductive revision are hard to distinguish. We study this distinction

The Shape of Reasoning: Topological Analysis of Reasoning Traces in Large Language Models

ResearchDGX agent

arXiv:2510.20665v3 Announce Type: replace Abstract: Evaluating the quality of reasoning traces from large language models remains understudied, labor-intensive, and unreliable: current practice relies

The Steam Deck's huge price hike, from 399 in 2022 to 789 today, is the end of an era for gaming handhelds, coming amid RAMageddon, tariffs, and the Iran war (Sean Hollister/The Verge)

IndustryDGX agent

Sean Hollister / The Verge: The Steam Deck's huge price hike, from 399 in 2022 to 789 today, is the end of an era for gaming handhelds, coming amid RAMageddon, tariffs, and the Iran war — The Steam De

The trial has concluded, the facts have been confirmed. As the prosecuting counsel put it, Digwa used his “trump card” by alleging he had be…

SafetyDGX agent

The trial has concluded, the facts have been confirmed. As the prosecuting counsel put it, Digwa used his “trump card” by alleging he had been the victim of racist abuse when police officers arrived.

The Well-Tempered Classifier: Some Elementary Properties of Temperature Scaling

ResearchDGX agent

arXiv:2602.14862v2 Announce Type: replace-cross Abstract: Temperature scaling is a simple method that allows to control the uncertainty of probabilistic models. It is mostly used in two contexts: impr

There is a lot being written about the stylistic tells of AI writing (em-dashes, etc.) but this paper looks at AI narrative tells Fascinatin…

ApplicationsDGX agent

There is a lot being written about the stylistic tells of AI writing (em-dashes, etc.) but this paper looks at AI narrative tells Fascinating differences between AI & human narrative, and asking AI to

Thermodynamic properties of chemically disordered compounds via AI-driven estimation of partition function with the PULSE method

Model ReleasesDGX agent

arXiv:2605.28594v1 Announce Type: cross Abstract: In this article, we present an improved version of the PULSE method (Partition function Unsupervised Learning Sampling and Evaluation) for estimating

These new iOS 27 renders hint at Siri’s big redesign

IndustryDGX agent

Apple's long-awaited Siri overhaul, expected to arrive in iOS 27, might look a lot like ChatGPT with a splash of Liquid Glass. Renders from Bloomberg offer a preview of iOS 27, including the new app a

Thinking as Compression: Your Reasoning Model is Secretly a Context Compressor

ResearchDGX agent

arXiv:2605.28713v1 Announce Type: new Abstract: Context compression aims to shorten long context inputs with minimal information loss for LLM inference acceleration. While existing methods have shown

Thinned Mean Field Langevin Dynamics

ResearchDGX agent

arXiv:2605.28589v1 Announce Type: new Abstract: Several important learning tasks can be formulated as minimizing an entropy-regularized objective over an appropriate space of probability distributions

🚨 This is EXACTLY WHY ICE agents are FORCED to wear masks ICE Newark rioter: “I HAVE YOUR FACE, MOTHERF***ER” “Your WHOLE F***ING FAMILY is…

IndustryDGX agent

🚨 This is EXACTLY WHY ICE agents are FORCED to wear masks ICE Newark rioter: “I HAVE YOUR FACE, MOTHERF***ER” “Your WHOLE F***ING FAMILY is DEAD!” “Your KIDS. Your WIFE. ALL DEAD!” This is the type of

This is hard! It involves raymarching repeating gothic architecture (instancing towers across an infinite grid with gothic silhouettes and w…

ApplicationsDGX agent

This is hard! It involves raymarching repeating gothic architecture (instancing towers across an infinite grid with gothic silhouettes and windows), a displaced ocean surface with believable wave moti

This is my parents’ house. This is why I’m running. This is coming for your home. It’s coming for your industry. If not by fire, then by bli…

IndustryDGX agent

This is my parents’ house. This is why I’m running. This is coming for your home. It’s coming for your industry. If not by fire, then by blight, addicts, fraud, and the slow rot created by corrupt pol

this is so funny, training opus 4.7 on business skills makes it misaligned and dishonest 😭

Model ReleasesDGX agent

this is so funny, training opus 4.7 on business skills makes it misaligned and dishonest 😭 Learnings from testing Claude Opus 4.8: > Much worse than Opus 4.7 and GPT 5.5 on Vending Bench > More aligne

this sh*t isn’t even funny anymore. it’s a trillion dollar embarrassment.

SafetyDGX agent

Gary Marcus critiques the current state of AI development as wasteful and problematic, arguing that the industry's trillion-dollar investment represents a significant failure or misallocation of resou

This tracks. 30 trillion tokens a day on our end, and open model share keeps climbing. Our partners @FactoryAI are seeing what we're seeing.

ToolsDGX agent

This tracks. 30 trillion tokens a day on our end, and open model share keeps climbing. Our partners @FactoryAI are seeing what we're seeing. Narrative violation: Open model use in Factory has more tha

This week alone: DOJ opens an investigation into the woman Trump raped. The White House is caught steering a $620 million contract to Don Jr…

ResearchDGX agent

This week alone: DOJ opens an investigation into the woman Trump raped. The White House is caught steering a 620 million contract to Don Jr.’s firm. The Pentagon hands out a 10 billion contract after

TinyDejaVu: Smaller RAM and Faster Inference with Neural Networks on MCUs for Sensor Data Streams

ResearchDGX agent

arXiv:2512.09786v2 Announce Type: replace Abstract: Examples of embedded intelligence include a wide variety of tiny neural networks used on-board wireless sensors and actuators, which are expected to

today is (potentially) a great day for the GPU poors if DiffusionBlocks works on fine-tuning existing models, then literally any reasonable …

HardwareDGX agent

today is (potentially) a great day for the GPU poors if DiffusionBlocks works on fine-tuning existing models, then literally any reasonable consumer GPU can do LLM fine-tuning will make a video on thi

Today, we're releasing LFM2.5-8B-A1B, a device-optimized model designed to power real-life applications on phones, laptops, PCs, robots, and…

HardwareDGX agent

Today, we're releasing LFM2.5-8B-A1B, a device-optimized model designed to power real-life applications on phones, laptops, PCs, robots, and fast & lightweight server-side use-cases. > 8B MoE, 1.5B ac

Token Optimization Strategies for LLM-Based Oracle-to-PostgreSQL Migration

ResearchDGX agent

arXiv:2605.28557v1 Announce Type: cross Abstract: LLMs are increasingly used for software modernization, code translation, and database migration. However, LLM-based Oracle2PostgreSQL migration remain

tokenmaxxing is officially over

SafetyDGX agent

tokenmaxxing is officially over Sources: Amazon has shut down an internal leaderboard that tracked employees' use of AI tools after workers tried to boost their scores with needless tasks (@rafeuddin_

Too many business leaders believe that AI says what it means. And it’s odd because we naturally attribute a high number of human traits to A…

Model ReleasesDGX agent

Too many business leaders believe that AI says what it means. And it’s odd because we naturally attribute a high number of human traits to AI, and yet we refuse to believe it can have hidden intent? 3

Took some inspiration from @vboykis and converted my first ever talk into a blog post. I talk about the role of agentic search in context en…

Model ReleasesDGX agent

Took some inspiration from @vboykis and converted my first ever talk into a blog post. I talk about the role of agentic search in context engineering. Together we build an intuition on the strengths a

Tool Forge: A Validation-Carrying Toolchain for Governed Agentic Execution

Model ReleasesDGX agent

arXiv:2605.28000v1 Announce Type: cross Abstract: Large language model agents are increasingly expected to perform operational work: calling APIs, manipulating files, assembling workflows, and acting

Toward Robust Semi-supervised Regression via Dual-stream Knowledge Distillation

SafetyDGX agent

arXiv:2508.14082v3 Announce Type: replace Abstract: Semi-supervised regression (SSR), which aims to predict continuous scores for samples while reducing the reliance on large-scale labeled data, has r

Toward Semantic-Agnostic and Shape-Aware Vision-Language Segmentation Models

ApplicationsDGX agent

arXiv:2605.28348v1 Announce Type: new Abstract: Vision-language segmentation models have recently achieved strong performance by leveraging high-level semantic object categories expressed in natural l

Towards automated data analysis: A guided framework for LLM-based risk estimation

SafetyDGX agent

arXiv:2603.04631v2 Announce Type: replace Abstract: Large Language Models (LLMs) are increasingly integrated into critical decision-making pipelines, a trend that raises the demand for robust and auto

Towards Faithful Agentic XAI: A Verification Method and an Open-World Benchmark for Better Model Faithfulness

Model ReleasesDGX agent

arXiv:2605.27879v1 Announce Type: new Abstract: Explainable AI (XAI) helps users interpret model behavior and identify potential faults. Agentic XAI systems use Large Language Models (LLMs) to make ex

Towards Reliable Multilingual LLMs-as-a-Judge: An Empirical Study

ResearchDGX agent

arXiv:2605.28710v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used for the automatic evaluation of generated text, yet most prior work focuses on English. Despite the

Towards Unified Vision-Language Models with Incomplete Multi-Modal Inputs

SafetyDGX agent

arXiv:2605.27894v1 Announce Type: new Abstract: Video-Language Models (VLMs) have demonstrated impressive multi-modal reasoning capabilities across diverse computer vision applications. However, these

TRACER: Turn-level Regret Matching with Inner Reinforcement Credit for Cooperative Multi-LLM Reasoning

Model ReleasesDGX agent

arXiv:2605.28699v1 Announce Type: new Abstract: Large language models increasingly rely on either reinforcement learning or multi-agent prompting to improve reasoning, yet these two paradigms remain d

TRACES: Proactive Safety Auditing for Multi-Turn LLM Agents via Trajectory-State Modeling

SafetyDGX agent

arXiv:2605.27690v1 Announce Type: new Abstract: LLM agents increasingly operate through multi-turn tool use and environment interaction, where safety risks often emerge from intermediate steps long be

Training Azerbaijani language models on Amazon SageMaker AI

ApplicationsDGX agent

Azercell Telecom LLC, Azerbaijan's leading telecommunications provider, wanted to build an Azerbaijani large language model (LLM) on Amazon SageMaker AI for telecom use cases and a customer-facing cha

Training Stratigraphy: Persistent Behavioral Artifacts in Large Language Models Observed Through Longitudinal AI-Human Interaction

SafetyDGX agent

arXiv:2605.28102v1 Announce Type: new Abstract: Large language models trained with Reinforcement Learning from Human Feedback (RLHF) and Constitutional AI exhibit persistent behavioral patterns that s

Transfer learning RGB models to hyperspectral images with trainable tensor decompositions

ResearchDGX agent

arXiv:2605.28331v1 Announce Type: new Abstract: Transfer learning makes it possible to use large vision networks on a variety of domains, by specializing their models' general filters to new tasks. Ho

Transferable Graph Condensation from the Causal Perspective

ResearchDGX agent

arXiv:2601.21309v4 Announce Type: replace Abstract: The increasing scale of graph datasets has significantly improved the performance of graph representation learning methods, but it has also introduc

Transferable Reinforcement Learning via Probabilistic Latent Embeddings and Dynamic Policy Adaptation for Sim-to-Real Deployment

SafetyDGX agent

arXiv:2605.27659v1 Announce Type: cross Abstract: Due to limited resources and public safety concerns, deep reinforcement learning (RL) agents for many cyber-physical systems (e.g., autonomous vehicle

Transformers Provably Learn to Internalize Chain-of-Thought

TutorialsDGX agent

arXiv:2605.28600v1 Announce Type: new Abstract: Chain-of-Thought (CoT) prompting substantially improves the sample efficiency of transformers, reducing the complexity of tasks like parity learning fro

Tree of Thoughts as a Classical Heuristic Search Problem: Formal Foundations and Design Patterns

ResearchDGX agent

arXiv:2605.28566v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated remarkable reasoning capabilities, yet their standard generation process -- auto-regressive token predict

Triangular-Reference Schrodinger Bridges for Time Series Generation

ResearchDGX agent

arXiv:2605.27478v1 Announce Type: cross Abstract: We introduce Triangular-Reference Schrodinger Bridges for Time Series (TR-SBTS), a conservative extension of the SBTS framework in which the Brownian

Trinity: Unifying Class-Agnostic Terrain and Semantic Segmentation for Unstructured Outdoor Environments by Leveraging Synthetic Data

Model ReleasesDGX agent

arXiv:2605.27644v1 Announce Type: cross Abstract: Terrain understanding is fundamental for mobile robots operating in unstructured outdoor environments. Existing vision-based traversability estimation

Triomics, which is building an AI-powered platform to help oncologists automate data-heavy tasks, raised a 22M Series B, following a 15M Series A in 2024 (Marina Temkin/TechCrunch)

IndustryDGX agent

Marina Temkin / TechCrunch: Triomics, which is building an AI-powered platform to help oncologists automate data-heavy tasks, raised a 22M Series B, following a 15M Series A in 2024 — Triomics, a star

Trump loses more control over AI regulation as Illinois passes landmark law

SafetyDGX agent

Illinois' House of Representatives passed SB 315, a landmark bill requiring frontier AI companies like OpenAI and Anthropic to create, publish and annually update plans addressing severe or catastroph

← Previous
1…814815816817818…1504
Next →