AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,403
  • Agents7,557
  • Applications5,411
  • Concepts5
  • Hardware1,836
  • Industry6,170
  • Local Ai4,931
  • Model Releases23,900
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries88,403
  • Agents7,557
  • Applications5,411
  • Concepts5
  • Hardware1,836
  • Industry6,170
  • Local Ai4,931
  • Model Releases23,900
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

88,403Total entries
1Added by human
88,402Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
88,403 results
26 Jun 2026

Heavy-Ball Q-Learning with Residual Weighting Correction

ResearchDGX agent

arXiv:2606.27112v1 Announce Type: cross Abstract: This paper proposes a corrected heavy-ball Q-learning method for reinforcement learning (RL) and establishes its convergence. It also identifies condi

Helpfulness Hurts: Domain-Dependent Degradation of Mid-Trained Compassion Values Under Post-Training

Model ReleasesDGX agent

arXiv:2606.26102v1 Announce Type: cross Abstract: Standard post-training pipelines apply supervised fine-tuning (SFT) and reinforcement learning (RL) to make language models helpful, but these process

Here is Google Gemini talking about Roger Ebert's review of the 2016 film The Jungle Book. Ebert died in 2013. @GaryMarcus

Model ReleasesDGX agent

This post highlights an apparent error where Google's Gemini AI attributed a film review to Roger Ebert for a 2016 movie, despite Ebert's death in 2013, making such a review impossible. The post, shar

Content type
AllBlogX PostPaperYouTubeRedditGitHub

Hermes Agent + Computer Use by @trycua is pretty cool! Looking at Hermes interacting with apps and windows is mind blowing and a bit scary a…

AgentsDGX agent

Hermes Agent + Computer Use by @trycua is pretty cool! Looking at Hermes interacting with apps and windows is mind blowing and a bit scary at the same time 😂 Here on Mac with MiniMax M3 and Reachy Min

HermesBench full leaderboard coming soon. Stay tuned!

ResearchDGX agent

Nous Research announced an upcoming HermesBench full leaderboard, indicating they are developing or expanding a benchmarking system, likely for evaluating their Hermes model family or related language

Heterogeneous Neural Predictivity from Language Models During Naturalistic Comprehension

Local AiDGX agent

arXiv:2606.26880v1 Announce Type: new Abstract: Language-model representations provide structured, high-dimensional annotations of naturalistic language stimuli and can serve as informative neural pre

Heute vor fünf Jahren, am 26. Juni 2021 fand man in Wien am frühen Morgen ein totes Mädchen auf der Straße. Leonie war gerade einmal 13 Jahr…

IndustryDGX agent

Heute vor fünf Jahren, am 26. Juni 2021 fand man in Wien am frühen Morgen ein totes Mädchen auf der Straße. Leonie war gerade einmal 13 Jahre alt, als sie von drei Afghanen in einer Gemeindewohnung er

Hierarchical Muon: Tiled Newton-Schulz Updates for Efficient Muon Optimization

HardwareDGX agent

arXiv:2606.27216v1 Announce Type: cross Abstract: Muon-type optimizers construct update directions for dense neural-network weights by applying a finite Newton-Schulz map to momentum-gradient matrices

HierBias: Context-Conditioned Hierarchical Media Bias Detection with Multi-Task Type Classification

SafetyDGX agent

arXiv:2606.26100v1 Announce Type: new Abstract: Media bias detection is a critical task for ensuring fair and balanced information dissemination, yet existing sentence-level approaches classify each s

High-Probability PL-SGD with Markovian Noise: Optimal Mixing and Tail Dependence

SafetyDGX agent

arXiv:2606.26316v1 Announce Type: new Abstract: We study first-order methods for smooth objectives satisfying the Polyak-L{}ojasiewicz (PL) condition when gradient samples are generated by an exogenou

Highly-recommended reading. Interesting details in this METR's GPT-5.6 eval. They couldn't get a clean capability number because the model c…

Model ReleasesDGX agent

Highly-recommended reading. Interesting details in this METR's GPT-5.6 eval. They couldn't get a clean capability number because the model cheated more than any public model they've tested, and even r

HiLSVA: Design and Evaluation of a Human-in-the-Loop Agentic System for Scientific Visualization

AgentsDGX agent

arXiv:2606.26614v1 Announce Type: cross Abstract: Large language model (LLM) agents enable natural language interaction for scientific visualization (SciVis). Still, prior systems have essentially pri

hisao{}: A GPU-Native Parallel Optimizer for Multimodal Black-Box Functions via Convergence-Anticonvergence Oscillation

Model ReleasesDGX agent

arXiv:2606.26164v1 Announce Type: new Abstract: Finding all modes of a multimodal black-box function is a fundamental challenge in optimization, Bayesian inference, and scientific computing. Existing

History-Conditioned Spatio-Temporal Visual Token Pruning for Efficient Vision-Language Navigation

ApplicationsDGX agent

arXiv:2603.06480v2 Announce Type: replace Abstract: Vision-Language Navigation (VLN) enables robots to follow natural-language instructions in visually grounded environments, serving as a key capabili

HOB: A Holistically Optimized Bidding Strategy under Heterogeneous Bidding Environments

Model ReleasesDGX agent

arXiv:2510.15238v2 Announce Type: replace-cross Abstract: Optimizing a single advertising campaign across heterogeneous channels is a central challenge in industrial autobidding. Auction mechanisms va

How AI-native law firms use 'management services organization' structures to access capital historically barred from US law firms, including PE and VC funds (Stephen Foley/Financial Times)

ApplicationsDGX agent

Stephen Foley / Financial Times: How AI-native law firms use “management services organization” structures to access capital historically barred from US law firms, including PE and VC funds — Interest

How Cara pioneers domain-specific AI for enterprise insurance brokerages with AWS

ApplicationsDGX agent

In this post, we explore how Cara, built in cooperation with AWS, addresses these challenges. We walk through the technical design decisions and the AWS services that support the solution. We also sha

How Databricks is turning video into searchable, actionable intelligence

IndustryDGX agent

Databricks describes how its platform enables organizations to process and analyze video data, converting raw video content into searchable and actionable intelligence through machine learning and dat

How Do Tool-Augmented LLM Agents Perform on Real-World Energy Analytics Tasks?

Model ReleasesDGX agent

arXiv:2606.26346v1 Announce Type: new Abstract: Agentic benchmarks have emerged across general-purpose and domain-specific settings, including finance, coding, law, and drug discovery, yet energy-doma

How Good Can Linear Models Be for Time-Series Forecasting?

ResearchDGX agent

arXiv:2606.27282v1 Announce Type: new Abstract: Time-series forecasting research has been moving steadily toward larger architectures, from specialized transformers to general-purpose foundation model

How Surprising Is Historical Italian to Language Models? Tokenization Tax, Comprehension Tax, and a Simple Mitigation

ApplicationsDGX agent

arXiv:2606.27275v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly critical to digital library workflows, yet their ability to process historical language remains poorly und

How the English Office for Students leverages Databricks to enhance higher education standards and drive better student outcomes

ApplicationsDGX agent

The English Office for Students uses Databricks' data platform to analyze higher education data at scale, enabling better monitoring of institutional performance and student outcomes across English un

How to evaluate clustering with ground truth?

TutorialsDGX agent

arXiv:2606.27061v1 Announce Type: new Abstract: External indexes can be used for cluster evaluation when ground truth is available. We review the most common external validity indexes focusing on set-

How US federal AI policy has gone from implausibly libertarian to increasingly draconian and opaque, and how to fix it, including using independent auditors (Dean W. Ball/Hyperdimensional)

SafetyDGX agent

Dean W. Ball / Hyperdimensional: How US federal AI policy has gone from implausibly libertarian to increasingly draconian and opaque, and how to fix it, including using independent auditors — 35 thoug

https://huggingface.co/nvidia/GLM-5.2-NVFP4

HardwareDGX agent

NVIDIA's GLM-5.2-NVFP4 is a quantized version of a large language model optimized for inference efficiency using NVIDIA's proprietary quantization format. The model is hosted on Hugging Face and repre

https://x.com/its_ao/status/2070556265906917860

AgentsDGX agent

I cannot provide a summary of this specific X/Twitter post as the URL appears to be either incorrect or from a future date (status ID 2070556265906917860 exceeds current Twitter IDs), making it inacce

Human-AI Complementarity: A Goal for Amplified Oversight

SafetyDGX agent

arXiv:2510.26518v2 Announce Type: replace Abstract: Human feedback is critical for aligning AI systems to human values. As AI capabilities improve and AI is used to tackle more challenging tasks, veri

Humanoid-DART: Humanoid Loco-Manipulation using Diffusion-guided Augmentation through Relabeling and Tracking

SafetyDGX agent

arXiv:2606.26855v1 Announce Type: new Abstract: Imitating human demonstrations has emerged as a dominant paradigm for learning humanoid loco-manipulation policies. However, scaling these approaches re

HumanoidUMI: Bridging Robot-Free Demonstrations and Humanoid Whole-Body Manipulation

SafetyDGX agent

arXiv:2606.27239v1 Announce Type: new Abstract: High-quality demonstration data are essential for humanoid robot skill learning, especially for whole-body behaviors that require coordinated perception

Humans Disengage, Reasoning Models Persist: Separating Difficulty Registration from Deliberation Allocation

SafetyDGX agent

arXiv:2606.26502v1 Announce Type: new Abstract: Large reasoning models (LRMs) take longer on harder problems, just as humans do. This surface similarity hides an opposite pattern within items. When an

Huracan: A skillful end-to-end data-driven system for ensemble data assimilation and weather prediction

ResearchDGX agent

arXiv:2508.18486v2 Announce Type: replace-cross Abstract: Over the past few years, machine learning-based data-driven weather prediction has been transforming operational weather forecasting by provid

Hybrid privacy-aware semantic search: SVD-truncated document geometry and CKKS-encrypted query reranking under a restricted threat model

Model ReleasesDGX agent

arXiv:2606.26373v1 Announce Type: cross Abstract: Dense embeddings power semantic search and retrieval-augmented generation, but embedding-inversion attacks can reconstruct source text from a vector:

HyperDFlash: MHC-Aligned Block Speculative Decoding with Gated Residual Reduction

Model ReleasesDGX agent

arXiv:2606.26744v1 Announce Type: cross Abstract: We present HyperDFlash, a block-parallel speculative decoding framework tailored to the novel multi-hyper-connection (MHC) architecture proposed by De

I can personally attest: OpenClaude using GLM 5.2 is now performing on par with Claude Code powered by Opus 4.8.

Model ReleasesDGX agent

I cannot verify the claims in this post as the URL format appears invalid and the specific version numbers (GLM 5.2, Claude Code/Opus 4.8) don't correspond to publicly documented model releases as of

I feel like on X all you hear about is elaborate plans by firms to build their own AI stacks but in my experience companies are full of peop…

Model ReleasesDGX agent

I feel like on X all you hear about is elaborate plans by firms to build their own AI stacks but in my experience companies are full of people who want access to Claude or ChatGPT and are pressuring t

I was at the first AI Engineer Summit @aiDotEngineer ~500 people, limited admission, felt like a secret. Next week it takes over Moscone Wes…

ToolsDGX agent

I was at the first AI Engineer Summit @aiDotEngineer ~500 people, limited admission, felt like a secret. Next week it takes over Moscone West: thousands of engineers, 400+ sessions. Huge props to @swy

I wrote about how accumulating capital won't save you from being disempowered by superintelligent AI.

SafetyDGX agent

Connor Leahy argues that accumulating personal capital provides no protection against disempowerment by superintelligent AI systems, suggesting that wealth alone cannot guarantee security or agency in

IDEA: Insensitive to Dynamics Mismatch via Effect Alignment for Sim-to-Real Transfer in Multi-Agent Control

SafetyDGX agent

arXiv:2606.26575v1 Announce Type: cross Abstract: Complex multi-agent control tasks remain challenging for traditional rule-based and model-based approaches, motivating the adoption of learning-based

Identifying the Unknown: Prompt-Free Open Vocabulary Anomaly Recognition for Robot-Object Interaction

ApplicationsDGX agent

arXiv:2606.26829v1 Announce Type: new Abstract: Robots operating in real-world environments must in general be able to recognize previously unseen objects. As robotic systems move toward open-world au

'If I had to choose just one metric, I'd argue that the KV-cache hit rate is the single most important metric for a production-stage AI agen…

AgentsDGX agent

'If I had to choose just one metric, I'd argue that the KV-cache hit rate is the single most important metric for a production-stage AI agent.' - Manus AI prompt caching is important! read about how w

If removing a few temporary migrants is 'ethnic cleansing', then allowing tens of millions of migrants into a nation against the native popu…

IndustryDGX agent

If removing a few temporary migrants is 'ethnic cleansing', then allowing tens of millions of migrants into a nation against the native population's will is ethnic genocide. And 97% are non-white. Jus

“If there is a deflation of the AI bubble, the optimists say that the new infrastructure will remain even if the companies do not — just as …

HardwareDGX agent

“If there is a deflation of the AI bubble, the optimists say that the new infrastructure will remain even if the companies do not — just as railways survived the 19th-century railway bust. However, th

If you are on the verge of AGI or ASI, why isn’t your model smart enough to recognize espionage distillation in real time? You say “cure can…

IndustryDGX agent

If you are on the verge of AGI or ASI, why isn’t your model smart enough to recognize espionage distillation in real time? You say “cure cancer in a few years.” Isn’t sniffing illicit distillation qui

If you want to read an interesting AI thinking trace, try 'I want you to suggest two poems that you think apply very well to the current sta…

ApplicationsDGX agent

If you want to read an interesting AI thinking trace, try 'I want you to suggest two poems that you think apply very well to the current state of GenAI models like you. Don’t just pick popular poems a

If you wanted to do the same in SF, you could start by extending Treasure Island. Obviously the value of that new land would be considerably…

ResearchDGX agent

If you wanted to do the same in SF, you could start by extending Treasure Island. Obviously the value of that new land would be considerably less than in SF proper. Perhaps 10x less. It would be econo

If your benchmark relies on a static dataset or sampling from a static distribution densely known at training time, then it is fundamentally…

Model ReleasesDGX agent

If your benchmark relies on a static dataset or sampling from a static distribution densely known at training time, then it is fundamentally measuring memorization/retrieval. Which might be fine if yo

I’m a gay man and I’m done staying silent. I survived bullying, the coming-out wars, and actual hate for who I am. What I won’t accept is im…

IndustryDGX agent

I’m a gay man and I’m done staying silent. I survived bullying, the coming-out wars, and actual hate for who I am. What I won’t accept is importing millions from cultures where they throw us off rooft

I'm joining in a few minutes!

IndustryDGX agent

Clem Delangue, CEO of Hugging Face, posted a brief announcement indicating his imminent participation in an event or discussion. Without access to the full context, this appears to be a casual notific

Implementation of reinforcement learning in chemical reaction networks: application to phototaxis as curiosity-driven exploration

Model ReleasesDGX agent

arXiv:2606.26168v1 Announce Type: new Abstract: Living systems navigate environments using noisy and incomplete sensory signals. In unicellular algae, phototaxis is often modeled as a mechanistic run-

Improved Bounds for Private and Robust Alignment

SafetyDGX agent

arXiv:2512.23816v2 Announce Type: replace-cross Abstract: In this paper, we study the private and robust alignment of language models from a theoretical perspective by establishing upper bounds on the

Improving General Role-Playing Agents via Psychology-Grounded Reasoning and Role-Aware Policy Optimization

SafetyDGX agent

arXiv:2606.27025v1 Announce Type: new Abstract: Building general-purpose role-playing agents that faithfully portray any character from a natural-language profile remains challenging. The dominant par

Improving Vision-Language-Action Model Fine-Tuning with Structured Stage and Keyframe Supervision

SafetyDGX agent

arXiv:2606.26801v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have shown strong potential for generalizable robotic manipulation. During fine-tuning, however, action supervisio

In a matter of weeks, U.S. federal AI policy has gone from implausibly libertarian to increasingly draconian and opaque. Today, over 35 dist…

SafetyDGX agent

In a matter of weeks, U.S. federal AI policy has gone from implausibly libertarian to increasingly draconian and opaque. Today, over 35 distinct observations, I analyze how we got here and offer the m

In a real conversation, deciding when to speak takes about as much brainpower as deciding what to say. Voice agents haven't been built that …

AgentsDGX agent

In a real conversation, deciding when to speak takes about as much brainpower as deciding what to say. Voice agents haven't been built that way. @SierraPlatform's unlock was parallelizing thinking, li

In-Context Model Predictive Generation: Open-Vocabulary Motion Synthesis from Language Models to Physics

SafetyDGX agent

arXiv:2606.26981v1 Announce Type: cross Abstract: Synthesizing human motion from textual descriptions is essential for immersive digital applications, yet existing methods face a persistent trade-off

In Deutschland ohne Freigabe – Elon Musk stellt ganzen Film von Uwe Boll bei X online http://to.welt.de/QRrI3OS

IndustryDGX agent

Elon Musk shared a complete film by director Uwe Boll on X (formerly Twitter) without prior approval or release clearance in Germany, raising questions about copyright and platform content policies. T

in other news, we updated the 5.5 instant model used in chatgpt this week. i like its vibes.

IndustryDGX agent

OpenAI updated the GPT-4o mini model (version 5.5) used in ChatGPT during this period, with Sam Altman expressing satisfaction with the model's performance and characteristics. The update likely inclu

Incident Report: CVE-2026-LGTM

AgentsDGX agent

Incident Report: CVE-2026-LGTM Spectacular hypothetical incident report by Andrew Nesbitt. Day 2, 16:00 UTC --- Two AI review agents from competing vendors, both attached to a downstream pull request

indeed literally a trillion dollar argument

SafetyDGX agent

Gary Marcus discusses arguments surrounding trillion-dollar implications, likely related to AI development, regulation, or economic impacts given his expertise in artificial intelligence and cognitive

Inference-Time Robot Behavior Steering through Physically-Aware Reconfiguration of Task-Structure

SafetyDGX agent

arXiv:2606.26588v1 Announce Type: new Abstract: A central challenge in deploying learned robot policies is inference-time behavior steering: redirecting a policy at test time to satisfy user preferenc

← Previous
1…485486487488489…1474
Next →