AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,419
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,934
  • Model Releases23,909
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,419
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,934
  • Model Releases23,909
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
Human
88,419Total entries
1Added by human
88,418Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
88,418 results
26 Jun 2026

Governing Actions, Not Agents: Institutional Attestation as a Governance Model for Autonomous AI Systems

SafetyDGX agent

arXiv:2606.26298v1 Announce Type: new Abstract: Autonomous AI agents may begin to perform consequential, irreversible actions such as clinical prescribing and production software deployment. This pape

GPT-5.6 Sol is our most capable model yet for cybersecurity. It shifts the performance-efficiency frontier for long-horizon security tasks i…

Model ReleasesDGX agent

GPT-5.6 Sol represents OpenAI's latest advancement in AI capabilities, specifically optimized for cybersecurity applications. The model demonstrates improved performance-efficiency tradeoffs, particul

GPT-5.6 Sol matches Mythos Preview on ExploitBench, adds Ultra mode with subagents for complex workflows, and max reasoning for deep problem-solving (OpenAI)

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model ReleasesDGX agent

OpenAI: GPT-5.6 Sol matches Mythos Preview on ExploitBench, adds Ultra mode with subagents for complex workflows, and max reasoning for deep problem-solving — We're beginning a limited preview of the

GPT‑5.6 Sol launches with our most robust safety stack yet. We strengthened real-time protections against high-risk cyber activity and repea…

Model ReleasesDGX agent

GPT‑5.6 Sol launches with our most robust safety stack yet. We strengthened real-time protections against high-risk cyber activity and repeated misuse, then spent weeks hardening the system with human

Gradient Testing and Estimation by Comparisons

ResearchDGX agent

arXiv:2405.11454v3 Announce Type: replace Abstract: We study gradient testing and gradient estimation of smooth functions using only a comparison oracle that, given two points, indicates which one has

Graph Neural Networks Applications Across Domains: All Insights You Need

SafetyDGX agent

arXiv:2606.27202v1 Announce Type: new Abstract: Graph neural networks have moved from a niche representation-learning technique to the default model class wherever data carry relational structure. The

Graph Reinforcement Learning for Calibration-Aware Quantum Circuit Routing

SafetyDGX agent

arXiv:2606.12816v3 Announce Type: replace-cross Abstract: Quantum circuit routing is a key step in compiling programs for noisy intermediate-scale quantum processors. Routes that appear efficient by s

Great experiment testing how good AIs are getting at very ambitious end-to-end coding tasks. Opus 4.7, in 14 hours, was able to build a soft…

ApplicationsDGX agent

Great experiment testing how good AIs are getting at very ambitious end-to-end coding tasks. Opus 4.7, in 14 hours, was able to build a software package that would take 2-17 weeks of human engineering

Great to see the new GPT-5.6 models finally announced. Sad to see this new release strategy where only a select few get access initially. No…

Model ReleasesDGX agent

Great to see the new GPT-5.6 models finally announced. Sad to see this new release strategy where only a select few get access initially. Not a win for our industry IMO. Open-source AI must win! Intro

Grok is balanced

IndustryDGX agent

Elon Musk claims that Grok, the AI assistant developed by his company xAI, maintains balanced perspectives and avoids excessive bias in its responses. The statement reflects Musk's positioning of Grok

Hallucination in World Models is Predictable and Preventable

Model ReleasesDGX agent

arXiv:2606.27326v1 Announce Type: cross Abstract: Modern generative world models render increasingly realistic action-controllable futures, yet they frequently hallucinate: rollouts remain visually fl

Hardware Design for Table Tennis Robot Capable of Beating Professional Players

SafetyDGX agent

arXiv:2606.26643v1 Announce Type: new Abstract: This paper focuses on the hardware specifications required for a table tennis robot to beat professional players. After analyzing the motions of elite p

HarmVideoBench: Benchmarking Harmful Video Understanding in Large Multimodal Models

Model ReleasesDGX agent

arXiv:2606.27187v1 Announce Type: cross Abstract: Large vision-language models (LVLMs) have recently shown immense potential in automated content moderation, sparking growing interest in developing ha

HauntAttack: When Attack Follows Reasoning as a Shadow

SafetyDGX agent

arXiv:2506.07031v5 Announce Type: replace-cross Abstract: Emerging Large Reasoning Models (LRMs) consistently excel in mathematical and reasoning tasks, showcasing remarkable capabilities. However, th

Have been taking different local open-weight LLMs for a test drive in different harnesses (Qwen-Code, Codex, Claude Code). 30B Mixture-of-Ex…

Model ReleasesDGX agent

Have been taking different local open-weight LLMs for a test drive in different harnesses (Qwen-Code, Codex, Claude Code). 30B Mixture-of-Expert models are kind of a nice sweet spot and can solve chal

Heat waves mess with your brain. Scientists are trying to figure out why.

ResearchDGX agent

It’s been hot in London this week. Really hot. A dangerous heat wave has hit Western Europe. Yesterday, the UK recorded its highest ever June temperature at 36.1 °C (about 97 °F). But as the weather a

Heavy-Ball Q-Learning with Residual Weighting Correction

ResearchDGX agent

arXiv:2606.27112v1 Announce Type: cross Abstract: This paper proposes a corrected heavy-ball Q-learning method for reinforcement learning (RL) and establishes its convergence. It also identifies condi

Helpfulness Hurts: Domain-Dependent Degradation of Mid-Trained Compassion Values Under Post-Training

Model ReleasesDGX agent

arXiv:2606.26102v1 Announce Type: cross Abstract: Standard post-training pipelines apply supervised fine-tuning (SFT) and reinforcement learning (RL) to make language models helpful, but these process

Here is Google Gemini talking about Roger Ebert's review of the 2016 film The Jungle Book. Ebert died in 2013. @GaryMarcus

Model ReleasesDGX agent

This post highlights an apparent error where Google's Gemini AI attributed a film review to Roger Ebert for a 2016 movie, despite Ebert's death in 2013, making such a review impossible. The post, shar

Hermes Agent + Computer Use by @trycua is pretty cool! Looking at Hermes interacting with apps and windows is mind blowing and a bit scary a…

AgentsDGX agent

Hermes Agent + Computer Use by @trycua is pretty cool! Looking at Hermes interacting with apps and windows is mind blowing and a bit scary at the same time 😂 Here on Mac with MiniMax M3 and Reachy Min

HermesBench full leaderboard coming soon. Stay tuned!

ResearchDGX agent

Nous Research announced an upcoming HermesBench full leaderboard, indicating they are developing or expanding a benchmarking system, likely for evaluating their Hermes model family or related language

Heterogeneous Neural Predictivity from Language Models During Naturalistic Comprehension

Local AiDGX agent

arXiv:2606.26880v1 Announce Type: new Abstract: Language-model representations provide structured, high-dimensional annotations of naturalistic language stimuli and can serve as informative neural pre

Heute vor fünf Jahren, am 26. Juni 2021 fand man in Wien am frühen Morgen ein totes Mädchen auf der Straße. Leonie war gerade einmal 13 Jahr…

IndustryDGX agent

Heute vor fünf Jahren, am 26. Juni 2021 fand man in Wien am frühen Morgen ein totes Mädchen auf der Straße. Leonie war gerade einmal 13 Jahre alt, als sie von drei Afghanen in einer Gemeindewohnung er

Hierarchical Muon: Tiled Newton-Schulz Updates for Efficient Muon Optimization

HardwareDGX agent

arXiv:2606.27216v1 Announce Type: cross Abstract: Muon-type optimizers construct update directions for dense neural-network weights by applying a finite Newton-Schulz map to momentum-gradient matrices

HierBias: Context-Conditioned Hierarchical Media Bias Detection with Multi-Task Type Classification

SafetyDGX agent

arXiv:2606.26100v1 Announce Type: new Abstract: Media bias detection is a critical task for ensuring fair and balanced information dissemination, yet existing sentence-level approaches classify each s

High-Probability PL-SGD with Markovian Noise: Optimal Mixing and Tail Dependence

SafetyDGX agent

arXiv:2606.26316v1 Announce Type: new Abstract: We study first-order methods for smooth objectives satisfying the Polyak-L{}ojasiewicz (PL) condition when gradient samples are generated by an exogenou

Highly-recommended reading. Interesting details in this METR's GPT-5.6 eval. They couldn't get a clean capability number because the model c…

Model ReleasesDGX agent

Highly-recommended reading. Interesting details in this METR's GPT-5.6 eval. They couldn't get a clean capability number because the model cheated more than any public model they've tested, and even r

HiLSVA: Design and Evaluation of a Human-in-the-Loop Agentic System for Scientific Visualization

AgentsDGX agent

arXiv:2606.26614v1 Announce Type: cross Abstract: Large language model (LLM) agents enable natural language interaction for scientific visualization (SciVis). Still, prior systems have essentially pri

hisao{}: A GPU-Native Parallel Optimizer for Multimodal Black-Box Functions via Convergence-Anticonvergence Oscillation

Model ReleasesDGX agent

arXiv:2606.26164v1 Announce Type: new Abstract: Finding all modes of a multimodal black-box function is a fundamental challenge in optimization, Bayesian inference, and scientific computing. Existing

History-Conditioned Spatio-Temporal Visual Token Pruning for Efficient Vision-Language Navigation

ApplicationsDGX agent

arXiv:2603.06480v2 Announce Type: replace Abstract: Vision-Language Navigation (VLN) enables robots to follow natural-language instructions in visually grounded environments, serving as a key capabili

HOB: A Holistically Optimized Bidding Strategy under Heterogeneous Bidding Environments

Model ReleasesDGX agent

arXiv:2510.15238v2 Announce Type: replace-cross Abstract: Optimizing a single advertising campaign across heterogeneous channels is a central challenge in industrial autobidding. Auction mechanisms va

How AI-native law firms use 'management services organization' structures to access capital historically barred from US law firms, including PE and VC funds (Stephen Foley/Financial Times)

ApplicationsDGX agent

Stephen Foley / Financial Times: How AI-native law firms use “management services organization” structures to access capital historically barred from US law firms, including PE and VC funds — Interest

How Cara pioneers domain-specific AI for enterprise insurance brokerages with AWS

ApplicationsDGX agent

In this post, we explore how Cara, built in cooperation with AWS, addresses these challenges. We walk through the technical design decisions and the AWS services that support the solution. We also sha

How Databricks is turning video into searchable, actionable intelligence

IndustryDGX agent

Databricks describes how its platform enables organizations to process and analyze video data, converting raw video content into searchable and actionable intelligence through machine learning and dat

How Do Tool-Augmented LLM Agents Perform on Real-World Energy Analytics Tasks?

Model ReleasesDGX agent

arXiv:2606.26346v1 Announce Type: new Abstract: Agentic benchmarks have emerged across general-purpose and domain-specific settings, including finance, coding, law, and drug discovery, yet energy-doma

How Good Can Linear Models Be for Time-Series Forecasting?

ResearchDGX agent

arXiv:2606.27282v1 Announce Type: new Abstract: Time-series forecasting research has been moving steadily toward larger architectures, from specialized transformers to general-purpose foundation model

How Surprising Is Historical Italian to Language Models? Tokenization Tax, Comprehension Tax, and a Simple Mitigation

ApplicationsDGX agent

arXiv:2606.27275v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly critical to digital library workflows, yet their ability to process historical language remains poorly und

How the English Office for Students leverages Databricks to enhance higher education standards and drive better student outcomes

ApplicationsDGX agent

The English Office for Students uses Databricks' data platform to analyze higher education data at scale, enabling better monitoring of institutional performance and student outcomes across English un

How to evaluate clustering with ground truth?

TutorialsDGX agent

arXiv:2606.27061v1 Announce Type: new Abstract: External indexes can be used for cluster evaluation when ground truth is available. We review the most common external validity indexes focusing on set-

How US federal AI policy has gone from implausibly libertarian to increasingly draconian and opaque, and how to fix it, including using independent auditors (Dean W. Ball/Hyperdimensional)

SafetyDGX agent

Dean W. Ball / Hyperdimensional: How US federal AI policy has gone from implausibly libertarian to increasingly draconian and opaque, and how to fix it, including using independent auditors — 35 thoug

https://huggingface.co/nvidia/GLM-5.2-NVFP4

HardwareDGX agent

NVIDIA's GLM-5.2-NVFP4 is a quantized version of a large language model optimized for inference efficiency using NVIDIA's proprietary quantization format. The model is hosted on Hugging Face and repre

https://x.com/its_ao/status/2070556265906917860

AgentsDGX agent

I cannot provide a summary of this specific X/Twitter post as the URL appears to be either incorrect or from a future date (status ID 2070556265906917860 exceeds current Twitter IDs), making it inacce

Human-AI Complementarity: A Goal for Amplified Oversight

SafetyDGX agent

arXiv:2510.26518v2 Announce Type: replace Abstract: Human feedback is critical for aligning AI systems to human values. As AI capabilities improve and AI is used to tackle more challenging tasks, veri

Humanoid-DART: Humanoid Loco-Manipulation using Diffusion-guided Augmentation through Relabeling and Tracking

SafetyDGX agent

arXiv:2606.26855v1 Announce Type: new Abstract: Imitating human demonstrations has emerged as a dominant paradigm for learning humanoid loco-manipulation policies. However, scaling these approaches re

HumanoidUMI: Bridging Robot-Free Demonstrations and Humanoid Whole-Body Manipulation

SafetyDGX agent

arXiv:2606.27239v1 Announce Type: new Abstract: High-quality demonstration data are essential for humanoid robot skill learning, especially for whole-body behaviors that require coordinated perception

Humans Disengage, Reasoning Models Persist: Separating Difficulty Registration from Deliberation Allocation

SafetyDGX agent

arXiv:2606.26502v1 Announce Type: new Abstract: Large reasoning models (LRMs) take longer on harder problems, just as humans do. This surface similarity hides an opposite pattern within items. When an

Huracan: A skillful end-to-end data-driven system for ensemble data assimilation and weather prediction

ResearchDGX agent

arXiv:2508.18486v2 Announce Type: replace-cross Abstract: Over the past few years, machine learning-based data-driven weather prediction has been transforming operational weather forecasting by provid

Hybrid privacy-aware semantic search: SVD-truncated document geometry and CKKS-encrypted query reranking under a restricted threat model

Model ReleasesDGX agent

arXiv:2606.26373v1 Announce Type: cross Abstract: Dense embeddings power semantic search and retrieval-augmented generation, but embedding-inversion attacks can reconstruct source text from a vector:

HyperDFlash: MHC-Aligned Block Speculative Decoding with Gated Residual Reduction

Model ReleasesDGX agent

arXiv:2606.26744v1 Announce Type: cross Abstract: We present HyperDFlash, a block-parallel speculative decoding framework tailored to the novel multi-hyper-connection (MHC) architecture proposed by De

I can personally attest: OpenClaude using GLM 5.2 is now performing on par with Claude Code powered by Opus 4.8.

Model ReleasesDGX agent

I cannot verify the claims in this post as the URL format appears invalid and the specific version numbers (GLM 5.2, Claude Code/Opus 4.8) don't correspond to publicly documented model releases as of

I feel like on X all you hear about is elaborate plans by firms to build their own AI stacks but in my experience companies are full of peop…

Model ReleasesDGX agent

I feel like on X all you hear about is elaborate plans by firms to build their own AI stacks but in my experience companies are full of people who want access to Claude or ChatGPT and are pressuring t

I was at the first AI Engineer Summit @aiDotEngineer ~500 people, limited admission, felt like a secret. Next week it takes over Moscone Wes…

ToolsDGX agent

I was at the first AI Engineer Summit @aiDotEngineer ~500 people, limited admission, felt like a secret. Next week it takes over Moscone West: thousands of engineers, 400+ sessions. Huge props to @swy

I wrote about how accumulating capital won't save you from being disempowered by superintelligent AI.

SafetyDGX agent

Connor Leahy argues that accumulating personal capital provides no protection against disempowerment by superintelligent AI systems, suggesting that wealth alone cannot guarantee security or agency in

IDEA: Insensitive to Dynamics Mismatch via Effect Alignment for Sim-to-Real Transfer in Multi-Agent Control

SafetyDGX agent

arXiv:2606.26575v1 Announce Type: cross Abstract: Complex multi-agent control tasks remain challenging for traditional rule-based and model-based approaches, motivating the adoption of learning-based

Identifying the Unknown: Prompt-Free Open Vocabulary Anomaly Recognition for Robot-Object Interaction

ApplicationsDGX agent

arXiv:2606.26829v1 Announce Type: new Abstract: Robots operating in real-world environments must in general be able to recognize previously unseen objects. As robotic systems move toward open-world au

'If I had to choose just one metric, I'd argue that the KV-cache hit rate is the single most important metric for a production-stage AI agen…

AgentsDGX agent

'If I had to choose just one metric, I'd argue that the KV-cache hit rate is the single most important metric for a production-stage AI agent.' - Manus AI prompt caching is important! read about how w

If removing a few temporary migrants is 'ethnic cleansing', then allowing tens of millions of migrants into a nation against the native popu…

IndustryDGX agent

If removing a few temporary migrants is 'ethnic cleansing', then allowing tens of millions of migrants into a nation against the native population's will is ethnic genocide. And 97% are non-white. Jus

“If there is a deflation of the AI bubble, the optimists say that the new infrastructure will remain even if the companies do not — just as …

HardwareDGX agent

“If there is a deflation of the AI bubble, the optimists say that the new infrastructure will remain even if the companies do not — just as railways survived the 19th-century railway bust. However, th

If you are on the verge of AGI or ASI, why isn’t your model smart enough to recognize espionage distillation in real time? You say “cure can…

IndustryDGX agent

If you are on the verge of AGI or ASI, why isn’t your model smart enough to recognize espionage distillation in real time? You say “cure cancer in a few years.” Isn’t sniffing illicit distillation qui

If you want to read an interesting AI thinking trace, try 'I want you to suggest two poems that you think apply very well to the current sta…

ApplicationsDGX agent

If you want to read an interesting AI thinking trace, try 'I want you to suggest two poems that you think apply very well to the current state of GenAI models like you. Don’t just pick popular poems a

← Previous
1…485486487488489…1474
Next →