AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,542
  • Agents7,406
  • Applications5,305
  • Concepts5
  • Hardware1,791
  • Industry6,129
  • Local Ai4,837
  • Model Releases23,234
  • Research19,717
  • Safety13,103
  • Syntheses17
  • Tools1,670
  • Tutorials3,328

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,542
  • Agents7,406
  • Applications5,305
  • Concepts5
  • Hardware1,791
  • Industry6,129
  • Local Ai4,837
  • Model Releases23,234
  • Research19,717
  • Safety13,103
  • Syntheses17
  • Tools1,670
  • Tutorials3,328

Source
HumanDGX agent

86,542Total entries
1Added by human
86,541Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,103 results
7 May 2026

Position: Agent Should Invoke External Tools ONLY When Epistemically Necessary

AgentsDGX agent

arXiv:2506.00886v3 Announce Type: replace Abstract: As large language models evolve into tool-augmented agents, a central question remains unresolved: when is external tool use actually justified? Exi

Power Distribution Bridges Sampling, Self-Reward RL, and Self-Distillation

Local AiDGX agent

arXiv:2605.04542v1 Announce Type: new Abstract: Recent analyses question whether reinforcement learning (RL) is responsible for strong reasoning in large language models (LLMs). At the same time, dist

Predict-then-Diffuse: Adaptive Response Length for Compute-Budgeted Inference in Diffusion LLMs

SafetyDGX agent

arXiv:2605.04215v1 Announce Type: new Abstract: Diffusion-based Large Language Models (D-LLMs) represent a promising frontier in generative AI, offering fully parallel token generation that can lead t

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Preference-Based Self-Distillation: Beyond KL Matching via Reward Regularization

SafetyDGX agent

arXiv:2605.05040v1 Announce Type: new Abstract: On-policy distillation is an efficient alternative to reinforcement learning, offering dense token-level training signals. However, its reliance on a st

Probabilistic Classification and Uncertainty Quantification of Sahara Desert Climate Using Feedforward Neural Networks

ResearchDGX agent

arXiv:2605.04286v1 Announce Type: new Abstract: Climate classification plays a vital role in agricultural planning, hydrological studies, and climate science. One of the most widely used systems for c

ProFit: Leveraging High-Value Signals in SFT via Probability-Guided Token Selection

SafetyDGX agent

arXiv:2601.09195v3 Announce Type: replace Abstract: Supervised fine-tuning (SFT) is a fundamental post-training strategy to align Large Language Models (LLMs) with human intent. However, traditional S

QKVShare: Quantized KV-Cache Handoff for Multi-Agent On-Device LLMs

Model ReleasesDGX agent

arXiv:2605.03884v1 Announce Type: new Abstract: Multi-agent LLM systems on edge devices need to hand off latent context efficiently, but the practical choices today are expensive re-prefill or full-pr

Quantum-inspired Reinforcement Learning for Synthesizable Drug Design

Model ReleasesDGX agent

arXiv:2409.09183v2 Announce Type: replace Abstract: Synthesizable molecular design (also known as synthesizable molecular optimization) is a fundamental problem in drug discovery, and involves designi

RetentiveKV: State-Space Memory for Uncertainty-Aware Multimodal KV Cache Eviction

ResearchDGX agent

arXiv:2605.04075v1 Announce Type: cross Abstract: Multimodal Large Language Models face severe challenges in computational efficiency and memory consumption due to the substantial expansion of the vis

Road Risk Monitor: A Deployable U.S. Road Incident Forecasting System with Live Weather and Road-Level Tiles

SafetyDGX agent

arXiv:2605.04242v1 Announce Type: new Abstract: Nationwide road-incident forecasting is a systems problem before it is a modeling problem. A usable service must connect historical incident archives, h

RoDyGS: Robust Dynamic Gaussian Splatting for Casual Videos

Model ReleasesDGX agent

arXiv:2412.03077v2 Announce Type: replace Abstract: 4D reconstruction from casually captured monocular videos is challenging due to inherent ambiguity in reconstructing dynamic 3D geometry. To address

Safety Must Precede the Deployment of Open-Ended AI

SafetyDGX agent

arXiv:2502.04512v3 Announce Type: replace Abstract: AI advancements have been significantly driven by a combination of foundation models and curiosity-driven learning aimed at increasing capability an

Same Voice, Different Lab: On the Homogenization of Frontier LLM Personalities

ResearchDGX agent

arXiv:2605.02897v1 Announce Type: cross Abstract: LLM assistant personalities play a critical role in user experience and perceived response quality. We present a large-scale experiment of frontier LL

“She said the theme of this party is the industrial age. And you came in dressed like a train wreck.” Asking AIs to think of the equivalent …

Model ReleasesDGX agent

“She said the theme of this party is the industrial age. And you came in dressed like a train wreck.” Asking AIs to think of the equivalent to this Hold Steady lyric, but for AI. Claude was the clear

Simplex rethinks software development with Codex

ApplicationsDGX agent

Simplex is an OpenAI initiative that reimagines software development processes by integrating Codex, OpenAI's AI code generation model, to enhance developer productivity and streamline coding workflow

SkCC: Portable and Secure Skill Compilation for Cross-Framework LLM Agents

Model ReleasesDGX agent

arXiv:2605.03353v1 Announce Type: cross Abstract: LLM-Agents have evolved into autonomous systems for complex task execution, with the SKILL.md specification emerging as a de facto standard for encaps

Smart Passive Acoustic Monitoring: Embedding a Classifier on AudioMoth Microcontroller

AgentsDGX agent

arXiv:2605.03412v1 Announce Type: cross Abstract: Passive Acoustic Monitoring (PAM) is an efficient and non-invasive method for surveying ecosystems at a reduced cost. Typically, autonomous recorders

Spotify launches Save to Spotify, a command-line tool that allows AI agents to upload AI-generated audio summaries and personal podcasts to a user's account (Terrence O'Brien/The Verge)

Model ReleasesDGX agent

Terrence O'Brien / The Verge: Spotify launches Save to Spotify, a command-line tool that allows AI agents to upload AI-generated audio summaries and personal podcasts to a user's account — A new comma

Stable Agentic Control: Tool-Mediated LLM Architecture for Autonomous Cyber Defense

Model ReleasesDGX agent

arXiv:2605.03034v1 Announce Type: new Abstract: Agentic systems involved in high-stake decision-making under adversarial pressure need formal guarantees not offered by existing approaches. Motivated b

Structured 3D Latents Are Surprisingly Powerful: Unleashing Generalizable Style with 2D Diffusion

ResearchDGX agent

arXiv:2605.04412v1 Announce Type: new Abstract: 3D asset generation plays a pivotal role in fields such as gaming and virtual reality, enabling the rapid synthesis of high-fidelity 3D objects from a s

SWAN: Semantic Watermarking with Abstract Meaning Representation

Model ReleasesDGX agent

arXiv:2605.04305v1 Announce Type: new Abstract: We introduce SWAN (Semantic Watermarking with Abstract Meaning Representation), a novel framework that embeds watermark signatures into the semantic str

Temporal Structure Matters for Efficient Test-Time Adaptation in Wearable Human Activity Recognition

SafetyDGX agent

arXiv:2605.04617v1 Announce Type: new Abstract: Wearable human activity recognition (WHAR) models often suffer from performance degradation under real-world cross-user distribution shifts. Test-time a

The Chrome extension expands what Codex can do for coding and work. From debugging browser flows to checking dashboards, conducting research…

Model ReleasesDGX agent

The Chrome extension expands what Codex can do for coding and work. From debugging browser flows to checking dashboards, conducting research, or updating CRMs, Codex can take on more of the tasks that

The most female-led product org in tech right now: Chief Product Officer: Ami Vora Claude Code/Cowork Head of Product: Cat Wu Claude Code/Co…

Model ReleasesDGX agent

The most female-led product org in tech right now: Chief Product Officer: Ami Vora Claude Code/Cowork Head of Product: Cat Wu Claude Code/Cowork Head of Eng: Fiona Fung Claude Platform Head of Product

Unified Framework of Distributional Regret in Multi-Armed Bandits and Reinforcement Learning

Model ReleasesDGX agent

arXiv:2605.05102v1 Announce Type: new Abstract: We study the distribution of regret in stochastic multi-armed bandits and episodic reinforcement learning through a unified framework. We formalize a di

Unsat Core Prediction through Polarity-Aware Representation Learning over Clause-Literal Hypergraphs

TutorialsDGX agent

arXiv:2605.04819v1 Announce Type: new Abstract: Graph neural networks have been widely used in Boolean satisfiability (SAT) tasks to learn structural information from SAT formulas. The goal of these s

U.S. intelligence says Iran can outlast Trump’s Hormuz blockade for months — via @washingtonpost https://www.washingtonpost.com/national-sec…

Model ReleasesDGX agent

U.S. intelligence says Iran can outlast Trump’s Hormuz blockade for months — via @washingtonpost https://www.washingtonpost.com/national-security/2026/05/07/cia-intelligence-iran-trump-blockade-missil

Vibe launches Dot, a wearable AI device for capturing real-world workplace conversations

Model ReleasesDGX agent

Vibe Inc., the creator of a contextual artificial intelligence workspace platform that handles real-world meetings, introduced a wearable device today that brings an AI assistant to professionals wher

VL-UniTrack: A Unified Framework with Visual-Language Prompts for UAV-Ground Visual Tracking

Model ReleasesDGX agent

arXiv:2605.04574v1 Announce Type: new Abstract: UAV-ground visual tracking (UGVT) aims to simultaneously track the same object from both the UAV and the ground view. However, existing two-stream metho

VTAgent: Agentic Keyframe Anchoring for Evidence-Aware Video TextVQA

Model ReleasesDGX agent

arXiv:2605.04870v1 Announce Type: new Abstract: Video text-based visual question answering (Video TextVQA) aims to answer questions by reasoning over visual textual content appearing in videos. Despit

We know you’re eager for voice updates in ChatGPT. Stay tuned, we’re cooking.

Model ReleasesDGX agent

OpenAI announced on X that voice feature updates for ChatGPT are in development and coming soon, though no specific timeline was provided. The post acknowledges user demand for enhanced voice capabili

Welcome to DS4, a specialized inference engine for DeepSeek v4 Flash. https://github.com/antirez/ds4 This project would have been impossible…

Model ReleasesDGX agent

Welcome to DS4, a specialized inference engine for DeepSeek v4 Flash. https://github.com/antirez/ds4 This project would have been impossible without the existence of llama.cpp and GGML and the work of

What an amazing event! Thank you @AnthropicAI. Doctors’ coding community in full swing.

Model ReleasesDGX agent

Boris Cherny praised an Anthropic AI event that brought together a doctors' coding community, expressing enthusiasm about the gathering's energy and engagement. The post suggests Anthropic hosted or s

While Anthropic will use the Colossus 1 data center, which has a really bad environmental record, xAI retains the larger Colossus 2 for its own AI training (Simon Willison/Simon Willison's Weblog)

Model ReleasesDGX agent

Simon Willison / Simon Willison's Weblog: While Anthropic will use the Colossus 1 data center, which has a really bad environmental record, xAI retains the larger Colossus 2 for its own AI training —

With the help of Claude Mythos Preview, the Firefox team fixed more security bugs in April than in the past 15 months combined.

Model ReleasesDGX agent

The Firefox development team reportedly resolved a significantly higher volume of security vulnerabilities in April with assistance from Claude Mythos Preview, an AI tool, compared to their bug-fixing

Wordle 1,782 3/6 ⬛⬛⬛⬛🟨 🟩🟩🟨⬛⬛ 🟩🟩🟩🟩🟩

Model ReleasesDGX agent

This post shows a Wordle game result where the player solved puzzle #1,782 in 3 attempts, displaying the emoji grid that represents their guesses and letter placements (gray for incorrect letters, yel

Wordle 1,783 6/6 ⬛⬛⬛⬛🟩 ⬛⬛⬛🟩⬛ ⬛⬛🟩🟩🟩 ⬛🟩🟩🟩🟩 ⬛🟩🟩🟩🟩 🟩🟩🟩🟩🟩

Model ReleasesDGX agent

This post documents a Wordle game completion (puzzle #1,783) where the player solved the word in six attempts, the maximum allowed before failure. The color-coded emoji grid shows the progression of g

6 May 2026

A critical question in agent design is “how do we build agentic workflows so humans are given significant, interesting, or variance-producin…

Model ReleasesDGX agent

A critical question in agent design is “how do we build agentic workflows so humans are given significant, interesting, or variance-producing decisions as they come up in the work?” A Claude-run compa

A Deeper Dive into the Irreversibility of PolyProtect: Making Protected Face Templates Harder to Invert

Model ReleasesDGX agent

arXiv:2605.03857v1 Announce Type: new Abstract: This work presents a deeper analysis of the 'irreversibility' property of PolyProtect, a biometric template protection method initially proposed for sec

Adaptive Negative Scheduling for Graph Contrastive Learning

Model ReleasesDGX agent

arXiv:2605.03076v1 Announce Type: new Abstract: Graph contrastive learning (GCL) has become a central paradigm for self-supervised representation learning in computational intelligence, with applicati

AI Expert Twin: Capturing Expert Cognition for Human-Centred, Practice-Based Learning

ApplicationsDGX agent

arXiv:2605.01401v1 Announce Type: cross Abstract: Tacit knowledge embedded in expert practice remains difficult to capture, formalise, and scale. While AI-driven educational systems have advanced pers

AI-Generated Smells: An Analysis of Code and Architecture in LLM and Agent-Driven Development

AgentsDGX agent

arXiv:2605.02741v1 Announce Type: cross Abstract: The promise of Large Language Models in automated software engineering is often measured by functional correctness, overlooking the critical issue of

[AINews] Silicon Valley gets Serious about Services

HardwareDGX agent

Silicon Valley companies are increasingly shifting focus toward AI service offerings and applications rather than solely developing foundational models, reflecting a maturing market where practical de

Anthropic to use SpaceX’s Colossus 1 supercomputer for inference

Model ReleasesDGX agent

Anthropic PBC today announced that it will use SpaceX Corp.’s Colossus 1 supercomputer to power its Claude chatbot. The system was originally built in 2024 by xAI Holdings Corp., an artificial intelli

AsymK-Talker: Real-Time and Long-Horizon Talking Head Generation via Asymmetric Kernel Distillation

ResearchDGX agent

arXiv:2605.02948v1 Announce Type: new Abstract: Recent advances in diffusion models have markedly enhanced the visual fidelity of audio-driven talking head generation. Nevertheless, existing methods a

Bandits attack function optimization

Model ReleasesDGX agent

arXiv:2605.03496v1 Announce Type: new Abstract: We consider function optimization as a sequential decision making problem under budget constraint. This constraint limits the number of objective functi

Beyond State Machines: Executing Network Procedures with Agentic Tool-Calling Sequences

AgentsDGX agent

arXiv:2605.02584v1 Announce Type: cross Abstract: Agentic AI will be an essential enabling technology for designing future mobile communication systems, which could provide flexible and customized ser

BFORE: Butterfly-Firefly Optimized Retinex Enhancement for Low-Light Image Quality Improvement

Model ReleasesDGX agent

arXiv:2605.03509v1 Announce Type: new Abstract: Low-light image enhancement is a fundamental challenge in computer vision and multimedia applications, as images captured under insufficient illuminatio

BIT.UA-AAUBS at ArchEHR-QA 2026: Evaluating Open-Source and Proprietary LLMs via Prompting in Low-Resource QA

Local AiDGX agent

arXiv:2605.03618v1 Announce Type: new Abstract: This paper presents the joint participation of the BIT.UA and AAUBS groups in the ArchEHR-QA 2026 shared task, which focuses on clinical question answer

Code with Claude is happening now! ▪︎ 9:00AM - Keynote ▪︎ 10:30AM - What's new in Claude Code ▪︎ 11:15AM - Building on Claude at GitHub scal…

Model ReleasesDGX agent

Code with Claude is happening now! ▪︎ 9:00AM - Keynote ▪︎ 10:30AM - What's new in Claude Code ▪︎ 11:15AM - Building on Claude at GitHub scale ▪︎ 12:00PM - Get to production faster with Managed Agents

Continuous-video agents (computer use, robotics, static scenes) burn compute re-ingesting pixels that didn't move. VLMaxxing teaches a froze…

Model ReleasesDGX agent

Continuous-video agents (computer use, robotics, static scenes) burn compute re-ingesting pixels that didn't move. VLMaxxing teaches a frozen video VLM to skip the reruns. 54 fps perception on Gemma 4

Deciphering Shortcut Learning from an Evolutionary Game Theory Perspective

SafetyDGX agent

arXiv:2605.02658v2 Announce Type: new Abstract: Shortcut learning causes deep learning models to rely on non-essential features within the data. However, its formation in deep neural network training

Delay, Plateau, or Collapse: Evaluating the Impact of Systematic Verification Error on RLVR

ApplicationsDGX agent

arXiv:2605.02909v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has become a powerful approach for improving the reasoning capabilities of large language models (

Distilling Long-CoT Reasoning through Collaborative Step-wise Multi-Teacher Decoding

ResearchDGX agent

arXiv:2605.02290v1 Announce Type: new Abstract: Distilling large reasoning models is essential for making Long-CoT reasoning practical, as full-scale inference remains computationally prohibitive. Exi

Do Not Waste Your Rollouts: Recycling Search Experience for Efficient Test-Time Scaling

ResearchDGX agent

arXiv:2601.21684v2 Announce Type: replace Abstract: Test-Time Scaling enhances the reasoning capabilities of Large Language Models by allocating additional inference compute to broaden the exploration

Elon has been the longest and by far the biggest voice actively warning about the dangers of AI for a very long time The entire reason he st…

Model ReleasesDGX agent

Elon has been the longest and by far the biggest voice actively warning about the dangers of AI for a very long time The entire reason he started OpenAI was this one thing.....to make sure AI is built

experiment: livetweeting the @AnthropicAI code with claude event! first up - @katelyn_lesse and @angjiang on claude platform!

Model ReleasesDGX agent

Swyx live-tweeted an Anthropic Code with Claude event, featuring speakers Katelyn Lesse and Ang Jiang discussing the Claude platform. The event appears to have been a real-time social media coverage o

Exploring Pass-Rate Reward in Reinforcement Learning for Code Generation

ResearchDGX agent

arXiv:2605.02944v1 Announce Type: new Abstract: Reinforcement learning (RL) from unit-test feedback has become a standard post-training recipe for improving large language models (LLMs) on code genera

FACTOR: Counterfactual Training-Free Test-Time Adaptation for Open-Vocabulary Object Detection

Model ReleasesDGX agent

arXiv:2605.03294v1 Announce Type: new Abstract: Open-vocabulary object detection often fails under distribution shifts, as it can be misled by spurious correlations between non-causal visual attribute

FluxFlow: Conservative Flow-Matching for Astronomical Image Super-Resolution

Model ReleasesDGX agent

arXiv:2605.03749v1 Announce Type: new Abstract: Ground-to-space astronomical super-resolution requires recovering space-quality images from ground-based observations that are simultaneously limited by

← Previous
1…691692693694695…1036
Next →