AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,993
  • Agents7,449
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,136
  • Local Ai4,859
  • Model Releases23,375
  • Research19,835
  • Safety13,176
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,993
  • Agents7,449
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,136
  • Local Ai4,859
  • Model Releases23,375
  • Research19,835
  • Safety13,176
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

86,993Total entries
1Added by human
86,992Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,477 results
28 May 2026

Privacy Protection Against Personalized Text-to-Image Synthesis via Cross-image Consistency Constraints

ResearchDGX agent

arXiv:2504.12747v2 Announce Type: replace Abstract: The rapid advancement of diffusion models and personalization techniques has made it possible to recreate individual portraits from just a few publi

RAG-Coding: Enhancing LLM Medical Coding with Structured External Knowledge

AgentsDGX agent

arXiv:2605.27377v1 Announce Type: cross Abstract: We present RAG-Coding, an agentic method for automated ICD-10-CM coding. RAG-Coding orchestrates four large language model (LLM) agents and grounds th

RAGe: A Retrieval-Augmented Generation Evaluation Framework

ResearchDGX agent

arXiv:2605.27445v1 Announce Type: cross Abstract: Deploying Large Language Model (LLM) applications, particularly those relying on Retrieval-Augmented Generation (RAG), remains challenging due to high

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

RASR: Retrieval-Augmented Super Resolution for Practical Reference-based Image Restoration

Model ReleasesDGX agent

arXiv:2508.09449v2 Announce Type: replace Abstract: Reference-based Super Resolution (RefSR) improves upon Single Image Super Resolution (SISR) by leveraging high-quality reference images to enhance t

REED: Post-Training Representation Editing for Cross-Domain Linguistic Steganalysis

Model ReleasesDGX agent

arXiv:2605.28298v1 Announce Type: new Abstract: In real-world scenarios of linguistic steganalysis, tested texts usually come from unseen domains with different vocabularies, topics, writing styles, a

Reevaluating Policy Gradient Methods for Imperfect-Information Games

Model ReleasesDGX agent

arXiv:2502.08938v4 Announce Type: replace Abstract: In the past decade, motivated by the putative failure of naive self-play deep reinforcement learning (DRL) in adversarial imperfect-information game

Refusal Before Decoding: Detecting and Exploiting Refusal Signals in Intermediate LLM Activations

SafetyDGX agent

arXiv:2605.28553v1 Announce Type: new Abstract: In this paper, we investigate whether refusal behavior can be predicted from LLM intermediate activations before decoding using linear probes trained on

Retrieval, Reward, and Training Protocols: What Matters in Training Search Agents?

SafetyDGX agent

arXiv:2605.27881v1 Announce Type: new Abstract: Search agents powered by large language models can autonomously decompose queries, retrieve information, and synthesize answers through multi-step reaso

ReverseMath: Answer Inversion for Scalable and Verifiable Mathematical Problem Generation

ResearchDGX agent

arXiv:2605.27709v1 Announce Type: new Abstract: Mathematical reasoning benchmarks are vital for evaluating large language models (LLMs), but many are static and repeatedly exposed through public evalu

RL Squeezes, SFT Expands: A Comparative Study of Reasoning LLMs

ResearchDGX agent

arXiv:2509.21128v2 Announce Type: replace Abstract: Large language models (LLMs) are typically trained by reinforcement learning (RL) with verifiable rewards (RLVR) and supervised fine-tuning (SFT) on

Robust Moment-Based Estimation via Spectral Gradient Reweighting

Model ReleasesDGX agent

arXiv:2605.27718v1 Announce Type: cross Abstract: Moment-based estimation is a theoretically attractive approach to parametric inference, especially when likelihood-based estimation is unavailable, mi

RW-TTT: Batched Serving for Request-Owned Test-Time Training State

Model ReleasesDGX agent

arXiv:2605.28053v1 Announce Type: new Abstract: Test-time training (TTT) adapts an LLM during generation by reading and updating request-owned state, such as fast weights, low-rank deltas, or streamin

Safe In-Context Reinforcement Learning

Model ReleasesDGX agent

arXiv:2509.25582v3 Announce Type: replace Abstract: In-context reinforcement learning (ICRL) is an emerging RL paradigm where an agent, after pretraining, can adapt to out-of-distribution test tasks w

Self-Consistency via Marginal Sharpening

ResearchDGX agent

arXiv:2605.28142v1 Announce Type: cross Abstract: Inference-time sampling can elicit strong reasoning abilities from language models without additional training. Existing power-sampling methods do so

Semiparametrically Efficient Inference for Kernel Measures of Noise Heterogeneity

SafetyDGX agent

arXiv:2605.27526v1 Announce Type: cross Abstract: We develop semiparametrically efficient inference for kernel measures of noise heterogeneity in additive noise models. In many applications, the regre

SHIPPED. Mistral Vibe is now the AI agent for long-horizon productivity and coding, and the home for Work mode, Code mode, the CLI, and a br…

Model ReleasesDGX agent

Mistral AI has released Mistral Vibe, an AI agent designed for long-horizon productivity and coding tasks, featuring Work mode, Code mode, a CLI, and additional capabilities. The product consolidates

SIGMA: Semantic-Difference Instruction-Grounding Mask Annotator for Text-Driven Image Manipulation Localization

Local AiDGX agent

arXiv:2605.27924v1 Announce Type: new Abstract: Text-driven image editing has advanced rapidly, but reliably localizing these manipulations requires image manipulation localization (IML) models traine

SkillGrad: Optimizing Agent Skills Like Gradient Descent

Model ReleasesDGX agent

arXiv:2605.27760v1 Announce Type: new Abstract: Agent skills provide a lightweight way to adapt LLM agents to specialized domains by storing reusable procedural knowledge in structured files. However,

Snowveil: A Framework for Decentralised Preference Discovery

Model ReleasesDGX agent

arXiv:2512.18444v2 Announce Type: replace-cross Abstract: Aggregating subjective preferences in social choice traditionally assumes a trusted central authority. In contrast, this paper formalises Dece

SSR3D-LLM: Structured Spatial Reasoning via Latent Steps for Fine-Grained Grounding in Unified 3D-LLMs

Model ReleasesDGX agent

arXiv:2605.28490v1 Announce Type: cross Abstract: 3D object grounding localizes referred objects in a 3D scene from natural language. Unified instance-centric 3D-LLMs aim to solve grounding together w

STR Robot: Design of an Autonomous Mobile Robot from Simulation to Reality

Model ReleasesDGX agent

arXiv:2605.28110v1 Announce Type: new Abstract: With the rapid development of simulation tools, the development and validation of autonomous robotic systems have become more efficient before real-worl

SuiChat-CN: Benchmarking Contextual Suicide Risk Assessment in Chinese Group Chats

Model ReleasesDGX agent

arXiv:2605.27911v1 Announce Type: new Abstract: Suicide is a critical global public health challenge, causing approximately 720,000 deaths each year and calling for timely, effective prevention strate

🚨super bad news for three of the biggest IPOs in history:

Model ReleasesDGX agent

🚨super bad news for three of the biggest IPOs in history: Companies are starting to question whether soaring AI spending is delivering meaningful returns. An AI consultant tells us a client recently s

Super excited to finally share Dynamic Workflows in Claude Code!! We built this a couple months ago, and it has slowly become a daily driver…

Model ReleasesDGX agent

Super excited to finally share Dynamic Workflows in Claude Code!! We built this a couple months ago, and it has slowly become a daily driver for a bunch of people at Anthropic. A few tips for getting

SYNAPSE: Neuro-Symbolic Visual Thought-to-Text Decoding via Topological Semantic Denoising

SafetyDGX agent

arXiv:2605.27790v1 Announce Type: new Abstract: Recent advances in large language models have accelerated open-vocabulary EEG-to-imagined-text decoding, where non-invasive neural activity recorded dur

TCP-MCP: Landscape-Guided Co-Evolution of Prompts and Communication Topologies for Multi-Agent Systems

Model ReleasesDGX agent

arXiv:2605.27850v1 Announce Type: new Abstract: Effective multi-agent systems cannot be designed by selecting prompts or communication graphs in isolation. Agent behavior depends on the information an

Techno Week

Model ReleasesDGX agent

Techno Week is a thematic event or content series from Cohere, a leading AI company, likely featuring updates on technological innovations, product announcements, or industry insights shared via their

Temporal Hyperbolic Graph Representation Learning for Scale-Free Internet Routing and Delay Prediction

ApplicationsDGX agent

arXiv:2605.28155v1 Announce Type: new Abstract: Predicting Internet round-trip time (RTT) is critical for routing optimization, quality-of-service (QoS) provisioning, and traffic engineering, yet rema

The Cases LJP Never Sees: Prosecution Decision Prediction for More Complete Criminal Liability Assessment

Model ReleasesDGX agent

arXiv:2605.28464v1 Announce Type: cross Abstract: Legal Judgment Prediction (LJP) has become a core benchmark for evaluating AI in the criminal legal domain, but it only sees criminal cases that have

The CFTC moves to vacate a $5M settlement with Gemini, reversing a Biden-era enforcement action, following a lobbying campaign by the Winklevoss twins (Wall Street Journal)

Model ReleasesDGX agent

Wall Street Journal: The CFTC moves to vacate a 5M settlement with Gemini, reversing a Biden-era enforcement action, following a lobbying campaign by the Winklevoss twins — A 5 million settlement at t

The European Commission launches a full review of JD.com's €2.2B acquisition of German electronics retailer Ceconomy under its Foreign Subsidies Regulation (Bloomberg)

Model ReleasesDGX agent

Bloomberg: The European Commission launches a full review of JD.com's €2.2B acquisition of German electronics retailer Ceconomy under its Foreign Subsidies Regulation — Chinese e-commerce firm JD.com

The Name’s Gaming … Cloud Gaming: ‘007 First Light’ Launches on GeForce NOW

Model ReleasesDGX agent

License to stream, shaken and stirred. GeForce NOW is dialing up the espionage with the launch of 007 First Light, letting members slip into James Bond’s reimagined origin story from almost any device

Too many business leaders believe that AI says what it means. And it’s odd because we naturally attribute a high number of human traits to A…

Model ReleasesDGX agent

Too many business leaders believe that AI says what it means. And it’s odd because we naturally attribute a high number of human traits to AI, and yet we refuse to believe it can have hidden intent? 3

Took some inspiration from @vboykis and converted my first ever talk into a blog post. I talk about the role of agentic search in context en…

Model ReleasesDGX agent

Took some inspiration from @vboykis and converted my first ever talk into a blog post. I talk about the role of agentic search in context engineering. Together we build an intuition on the strengths a

Towards automated data analysis: A guided framework for LLM-based risk estimation

SafetyDGX agent

arXiv:2603.04631v2 Announce Type: replace Abstract: Large Language Models (LLMs) are increasingly integrated into critical decision-making pipelines, a trend that raises the demand for robust and auto

Tree of Thoughts as a Classical Heuristic Search Problem: Formal Foundations and Design Patterns

ResearchDGX agent

arXiv:2605.28566v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated remarkable reasoning capabilities, yet their standard generation process -- auto-regressive token predict

Trinity: Unifying Class-Agnostic Terrain and Semantic Segmentation for Unstructured Outdoor Environments by Leveraging Synthetic Data

Model ReleasesDGX agent

arXiv:2605.27644v1 Announce Type: cross Abstract: Terrain understanding is fundamental for mobile robots operating in unstructured outdoor environments. Existing vision-based traversability estimation

UNIQUE: Universal Top-k Sparse Attention for Training-free Inference and Sparsity-aware Training

ResearchDGX agent

arXiv:2605.27740v1 Announce Type: new Abstract: Long-context inference in large language models (LLMs) is bottlenecked by the linear growth of the self-attention key-value (KV) cache. Top-k sparse att

Verified Misguidance: Measuring Structural Citation Failures in Search-Augmented LLMs

SafetyDGX agent

arXiv:2605.28565v1 Announce Type: cross Abstract: Users of search-augmented LLMs rely on citations as evidence that responses are grounded in real sources, and rarely verify the cited pages themselves

VeriTrip: A Verifiable Benchmark for Travel Planning Agents over Unstructured Web Corpora

Model ReleasesDGX agent

arXiv:2605.28683v1 Announce Type: new Abstract: Existing benchmarks have laid the foundation for travel planning agents by establishing API-centric paradigms. However, as the capabilities of Autonomou

wait… if most people think 5.5 is better than 4.7, i assume that’s due to terminal coding benchmark… 4.8 is still outperformed by 5.5

Model ReleasesDGX agent

wait… if most people think 5.5 is better than 4.7, i assume that’s due to terminal coding benchmark… 4.8 is still outperformed by 5.5 Introducing Claude Opus 4.8: it builds on Opus 4.7 with sharper ju

We also shipped dynamic workflows in Claude Code (research preview), for tasks too big for one pass. Make sure to default to auto mode so Cl…

Model ReleasesDGX agent

We also shipped dynamic workflows in Claude Code (research preview), for tasks too big for one pass. Make sure to default to auto mode so Claude isn't stopping for permissions. It's token-intensive, s

We've raised 65 billion in Series H funding at a 965 billion post-money valuation, led by @AltimeterCap, Dragoneer, @Greenoaks, and @sequo…

Model ReleasesDGX agent

We've raised 65 billion in Series H funding at a 965 billion post-money valuation, led by @AltimeterCap, Dragoneer, @Greenoaks, and @sequoia. This investment will help us advance our research and expa

When prompt perturbations break your A/B test: A valid statistical test for generative surveying

ResearchDGX agent

arXiv:2605.27463v1 Announce Type: cross Abstract: Generative surveying -- where collections of LLM-based personas provide feedback on messages -- has emerged as a cheap and scalable alternative to tra

Where Rollouts Begin: Low-Load, High-Leverage First-Token Diversification for RLVR

SafetyDGX agent

arXiv:2605.28295v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) trains reasoning models without labeled trajectories, relying on grouped rollouts to expose the po

Which Heads Matter for Reasoning? RL-Guided KV Cache Compression

ResearchDGX agent

arXiv:2510.08525v3 Announce Type: replace Abstract: Reasoning large language models exhibit complex reasoning behaviors via extended chain-of-thought generation that are highly fragile to information

Who's going first

Model ReleasesDGX agent

Who's going first Introducing Claude Opus 4.8: it builds on Opus 4.7 with sharper judgment, more honesty about its own progress, and the ability to work independently for longer than its predecessors.

Wordle 1,803 5/6 🟩🟩⬛⬛⬛ ⬛⬛⬛⬛⬛ 🟩🟩🟩⬛⬛ 🟩🟩🟩⬛⬛ 🟩🟩🟩🟩🟩

Model ReleasesDGX agent

This entry documents a Wordle game result (puzzle #1,803) where the player achieved a solution in 5 out of 6 allowed guesses. The color-coded emoji sequence shows the player's guess progression, with

You Live More Than Once: Towards Hierarchical Skill Meta-Evolving

Model ReleasesDGX agent

arXiv:2605.28390v1 Announce Type: new Abstract: Test-time skill evolving is regarded as a new paradigm for enhancing deployed agentic systems. Existing works mainly focus on hard-coded skill evolving

Zipping the Thought: When and How Compressed Reasoning Data Works in LLM Post-Training

ResearchDGX agent

arXiv:2605.28008v1 Announce Type: new Abstract: Large language models (LLMs) can now solve complex problems through long chain-of-thought (CoT) reasoning, but the trade-off between performance and tok

27 May 2026

A deep dive into how Anthropic's Claude Code and Peter Steinberger's OpenClaw unleashed the AI agent revolution that is rapidly transforming modern computing (Steven Levy/Wired)

Model ReleasesDGX agent

Steven Levy / Wired: A deep dive into how Anthropic's Claude Code and Peter Steinberger's OpenClaw unleashed the AI agent revolution that is rapidly transforming modern computing — The definitive stor

A multifractal-based masked auto-encoder: an application to medical images

ResearchDGX agent

arXiv:2605.26287v1 Announce Type: new Abstract: Masked autoencoders (MAE) have shown great promise in medical image classification. However, the random masking strategy employed by traditional MAEs ma

AGORA: Adapter-Grounded Observation-Action Retention for Inference-Free Prompt Compression in LLM Agents

Model ReleasesDGX agent

arXiv:2605.26596v1 Announce Type: new Abstract: The token-level extractive compressors widely used for general LM context are structurally inappropriate for LLM agents: across 17 (env, backbone, metho

AI evaluation may bias perceptions: The importance of context in interpreting academic writing

Model ReleasesDGX agent

arXiv:2605.26662v1 Announce Type: cross Abstract: This paper examines how estimates of AI use in scientific writing can be biased when evaluation methods ignore contextual differences across countries

Alignment Tampering: How Reinforcement Learning from Human Feedback Is Exploited to Optimize Misaligned Biases

SafetyDGX agent

arXiv:2605.27355v1 Announce Type: new Abstract: Reinforcement Learning from Human Feedback (RLHF) is the standard method to align Large Language Models (LLMs) with human preferences. In this work, we

Amazon MGM Studios announces the GenAI Creators' Fund, greenlights three AI animated series for Prime Video, and launches an AI production platform with AWS (Todd Spangler/Variety)

Model ReleasesDGX agent

Todd Spangler / Variety: Amazon MGM Studios announces the GenAI Creators' Fund, greenlights three AI animated series for Prime Video, and launches an AI production platform with AWS — Amazon MGM Studi

Annotator Positionality as Signal: Psychometric Weighting for Anti-Autistic Ableism Detection

SafetyDGX agent

arXiv:2605.26397v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used in decision-making tasks where they can amplify or suppress perspectives, raising concerns in high-

AutoDFT: A Closed-Loop Multi-Agent Framework for Autonomous DFT Calculations

Model ReleasesDGX agent

arXiv:2605.26179v1 Announce Type: cross Abstract: Density functional theory (DFT) serves as the basis for computational discovery in materials science and chemistry, yet each calculation demands exten

AWS launches Agentic Shopping Assistant to help retailers build AI tools

Model ReleasesDGX agent

Amazon Web Services Inc. today introduced a new offering designed to help retailers integrate artificial intelligence features into their online stores. AWS Agentic Shopping Assistant, or ASA, combine

BAIT: Boundary-Guided Disclosure Escalation via Self-Conditioned Reasoning

SafetyDGX agent

arXiv:2605.27110v1 Announce Type: cross Abstract: In this work, we propose BAIT (Boundary-Aware Iterative Trap), a three-step jailbreak framework that approaches malicious goals through internal discl

← Previous
1…662663664665666…1042
Next →