AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,678
  • Agents7,513
  • Applications5,367
  • Concepts5
  • Hardware1,821
  • Industry6,154
  • Local Ai4,902
  • Model Releases23,619
  • Research19,969
  • Safety13,271
  • Syntheses17
  • Tools1,674
  • Tutorials3,366

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,678
  • Agents7,513
  • Applications5,367
  • Concepts5
  • Hardware1,821
  • Industry6,154
  • Local Ai4,902
  • Model Releases23,619
  • Research19,969
  • Safety13,271
  • Syntheses17
  • Tools1,674
  • Tutorials3,366

Source
HumanDGX agent

Content type
AllBlog
87,678Total entries
1Added by human
87,677Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,033 results
Model Releases

The Cases LJP Never Sees: Prosecution Decision Prediction for More Complete Criminal Liability Assessment

DGX agent

arXiv:2605.28464v1 Announce Type: cross Abstract: Legal Judgment Prediction (LJP) has become a core benchmark for evaluating AI in the criminal legal domain, but it only sees criminal cases that have

model-releasesarxiv-cs-ai
28 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

The CFTC moves to vacate a $5M settlement with Gemini, reversing a Biden-era enforcement action, following a lobbying campaign by the Winklevoss twins (Wall Street Journal)

DGX agent

Wall Street Journal: The CFTC moves to vacate a 5M settlement with Gemini, reversing a Biden-era enforcement action, following a lobbying campaign by the Winklevoss twins — A 5 million settlement at t

model-releasestechmeme
28 May 2026
Model Releases

The European Commission launches a full review of JD.com's €2.2B acquisition of German electronics retailer Ceconomy under its Foreign Subsidies Regulation (Bloomberg)

DGX agent

Bloomberg: The European Commission launches a full review of JD.com's €2.2B acquisition of German electronics retailer Ceconomy under its Foreign Subsidies Regulation — Chinese e-commerce firm JD.com

model-releasestechmeme
28 May 2026
Model Releases

The Name’s Gaming … Cloud Gaming: ‘007 First Light’ Launches on GeForce NOW

DGX agent

License to stream, shaken and stirred. GeForce NOW is dialing up the espionage with the launch of 007 First Light, letting members slip into James Bond’s reimagined origin story from almost any device

model-releasesnvidia-blog
28 May 2026
Model Releases

Too many business leaders believe that AI says what it means. And it’s odd because we naturally attribute a high number of human traits to A…

DGX agent

Too many business leaders believe that AI says what it means. And it’s odd because we naturally attribute a high number of human traits to AI, and yet we refuse to believe it can have hidden intent? 3

model-releasesallie-k--miller--x
28 May 2026
Model Releases

Took some inspiration from @vboykis and converted my first ever talk into a blog post. I talk about the role of agentic search in context en…

DGX agent

Took some inspiration from @vboykis and converted my first ever talk into a blog post. I talk about the role of agentic search in context engineering. Together we build an intuition on the strengths a

model-releasesswyx--x
28 May 2026
Safety

Towards automated data analysis: A guided framework for LLM-based risk estimation

DGX agent

arXiv:2603.04631v2 Announce Type: replace Abstract: Large Language Models (LLMs) are increasingly integrated into critical decision-making pipelines, a trend that raises the demand for robust and auto

safetyarxiv-cs-ai
28 May 2026
Research

Tree of Thoughts as a Classical Heuristic Search Problem: Formal Foundations and Design Patterns

DGX agent

arXiv:2605.28566v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated remarkable reasoning capabilities, yet their standard generation process -- auto-regressive token predict

researcharxiv-cs-ai
28 May 2026
Model Releases

Trinity: Unifying Class-Agnostic Terrain and Semantic Segmentation for Unstructured Outdoor Environments by Leveraging Synthetic Data

DGX agent

arXiv:2605.27644v1 Announce Type: cross Abstract: Terrain understanding is fundamental for mobile robots operating in unstructured outdoor environments. Existing vision-based traversability estimation

model-releasesarxiv-cs-ai
28 May 2026
Research

UNIQUE: Universal Top-k Sparse Attention for Training-free Inference and Sparsity-aware Training

DGX agent

arXiv:2605.27740v1 Announce Type: new Abstract: Long-context inference in large language models (LLMs) is bottlenecked by the linear growth of the self-attention key-value (KV) cache. Top-k sparse att

researcharxiv-cs-cl
28 May 2026
Safety

Verified Misguidance: Measuring Structural Citation Failures in Search-Augmented LLMs

DGX agent

arXiv:2605.28565v1 Announce Type: cross Abstract: Users of search-augmented LLMs rely on citations as evidence that responses are grounded in real sources, and rarely verify the cited pages themselves

safetyarxiv-cs-ai
28 May 2026
Model Releases

VeriTrip: A Verifiable Benchmark for Travel Planning Agents over Unstructured Web Corpora

DGX agent

arXiv:2605.28683v1 Announce Type: new Abstract: Existing benchmarks have laid the foundation for travel planning agents by establishing API-centric paradigms. However, as the capabilities of Autonomou

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

wait… if most people think 5.5 is better than 4.7, i assume that’s due to terminal coding benchmark… 4.8 is still outperformed by 5.5

DGX agent

wait… if most people think 5.5 is better than 4.7, i assume that’s due to terminal coding benchmark… 4.8 is still outperformed by 5.5 Introducing Claude Opus 4.8: it builds on Opus 4.7 with sharper ju

model-releasesjeremy-howard--x
28 May 2026
Model Releases

We also shipped dynamic workflows in Claude Code (research preview), for tasks too big for one pass. Make sure to default to auto mode so Cl…

DGX agent

We also shipped dynamic workflows in Claude Code (research preview), for tasks too big for one pass. Make sure to default to auto mode so Claude isn't stopping for permissions. It's token-intensive, s

model-releasesboris-cherny--x
28 May 2026
Model Releases

We've raised 65 billion in Series H funding at a 965 billion post-money valuation, led by @AltimeterCap, Dragoneer, @Greenoaks, and @sequo…

DGX agent

We've raised 65 billion in Series H funding at a 965 billion post-money valuation, led by @AltimeterCap, Dragoneer, @Greenoaks, and @sequoia. This investment will help us advance our research and expa

model-releasessonya-huang--x
28 May 2026
Research

When prompt perturbations break your A/B test: A valid statistical test for generative surveying

DGX agent

arXiv:2605.27463v1 Announce Type: cross Abstract: Generative surveying -- where collections of LLM-based personas provide feedback on messages -- has emerged as a cheap and scalable alternative to tra

researcharxiv-cs-ai
28 May 2026
Safety

Where Rollouts Begin: Low-Load, High-Leverage First-Token Diversification for RLVR

DGX agent

arXiv:2605.28295v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) trains reasoning models without labeled trajectories, relying on grouped rollouts to expose the po

safetyarxiv-cs-ai
28 May 2026
Research

Which Heads Matter for Reasoning? RL-Guided KV Cache Compression

DGX agent

arXiv:2510.08525v3 Announce Type: replace Abstract: Reasoning large language models exhibit complex reasoning behaviors via extended chain-of-thought generation that are highly fragile to information

researcharxiv-cs-cl
28 May 2026
Model Releases

Who's going first

DGX agent

Who's going first Introducing Claude Opus 4.8: it builds on Opus 4.7 with sharper judgment, more honesty about its own progress, and the ability to work independently for longer than its predecessors.

model-releasesjerry-liu--x
28 May 2026
Model Releases

Wordle 1,803 5/6 🟩🟩⬛⬛⬛ ⬛⬛⬛⬛⬛ 🟩🟩🟩⬛⬛ 🟩🟩🟩⬛⬛ 🟩🟩🟩🟩🟩

DGX agent

This entry documents a Wordle game result (puzzle #1,803) where the player achieved a solution in 5 out of 6 allowed guesses. The color-coded emoji sequence shows the player's guess progression, with

model-releasesanthropic--x
28 May 2026
Model Releases

You Live More Than Once: Towards Hierarchical Skill Meta-Evolving

DGX agent

arXiv:2605.28390v1 Announce Type: new Abstract: Test-time skill evolving is regarded as a new paradigm for enhancing deployed agentic systems. Existing works mainly focus on hard-coded skill evolving

model-releasesarxiv-cs-ai
28 May 2026
Research

Zipping the Thought: When and How Compressed Reasoning Data Works in LLM Post-Training

DGX agent

arXiv:2605.28008v1 Announce Type: new Abstract: Large language models (LLMs) can now solve complex problems through long chain-of-thought (CoT) reasoning, but the trade-off between performance and tok

researcharxiv-cs-ai
28 May 2026
Model Releases

A deep dive into how Anthropic's Claude Code and Peter Steinberger's OpenClaw unleashed the AI agent revolution that is rapidly transforming modern computing (Steven Levy/Wired)

DGX agent

Steven Levy / Wired: A deep dive into how Anthropic's Claude Code and Peter Steinberger's OpenClaw unleashed the AI agent revolution that is rapidly transforming modern computing — The definitive stor

model-releasestechmeme
27 May 2026
Research

A multifractal-based masked auto-encoder: an application to medical images

DGX agent

arXiv:2605.26287v1 Announce Type: new Abstract: Masked autoencoders (MAE) have shown great promise in medical image classification. However, the random masking strategy employed by traditional MAEs ma

researcharxiv-cs-cv
27 May 2026
Model Releases

AGORA: Adapter-Grounded Observation-Action Retention for Inference-Free Prompt Compression in LLM Agents

DGX agent

arXiv:2605.26596v1 Announce Type: new Abstract: The token-level extractive compressors widely used for general LM context are structurally inappropriate for LLM agents: across 17 (env, backbone, metho

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

AI evaluation may bias perceptions: The importance of context in interpreting academic writing

DGX agent

arXiv:2605.26662v1 Announce Type: cross Abstract: This paper examines how estimates of AI use in scientific writing can be biased when evaluation methods ignore contextual differences across countries

model-releasesarxiv-cs-ai
27 May 2026
Safety

Alignment Tampering: How Reinforcement Learning from Human Feedback Is Exploited to Optimize Misaligned Biases

DGX agent

arXiv:2605.27355v1 Announce Type: new Abstract: Reinforcement Learning from Human Feedback (RLHF) is the standard method to align Large Language Models (LLMs) with human preferences. In this work, we

safetyarxiv-cs-ai
27 May 2026
Model Releases

Amazon MGM Studios announces the GenAI Creators' Fund, greenlights three AI animated series for Prime Video, and launches an AI production platform with AWS (Todd Spangler/Variety)

DGX agent

Todd Spangler / Variety: Amazon MGM Studios announces the GenAI Creators' Fund, greenlights three AI animated series for Prime Video, and launches an AI production platform with AWS — Amazon MGM Studi

model-releasestechmeme
27 May 2026
Safety

Annotator Positionality as Signal: Psychometric Weighting for Anti-Autistic Ableism Detection

DGX agent

arXiv:2605.26397v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used in decision-making tasks where they can amplify or suppress perspectives, raising concerns in high-

safetyarxiv-cs-ai
27 May 2026
Model Releases

AutoDFT: A Closed-Loop Multi-Agent Framework for Autonomous DFT Calculations

DGX agent

arXiv:2605.26179v1 Announce Type: cross Abstract: Density functional theory (DFT) serves as the basis for computational discovery in materials science and chemistry, yet each calculation demands exten

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

AWS launches Agentic Shopping Assistant to help retailers build AI tools

DGX agent

Amazon Web Services Inc. today introduced a new offering designed to help retailers integrate artificial intelligence features into their online stores. AWS Agentic Shopping Assistant, or ASA, combine

model-releasessiliconangle
27 May 2026
Safety

BAIT: Boundary-Guided Disclosure Escalation via Self-Conditioned Reasoning

DGX agent

arXiv:2605.27110v1 Announce Type: cross Abstract: In this work, we propose BAIT (Boundary-Aware Iterative Trap), a three-step jailbreak framework that approaches malicious goals through internal discl

safetyarxiv-cs-cl
27 May 2026
Model Releases

BEAT: Rhythm-Elastic Alignment for Agentic Music-guided Movie Trailer Generation

DGX agent

arXiv:2605.27067v1 Announce Type: new Abstract: Automatic movie trailer generation must select shots from a full-length film and synchronize them with background music. Existing methods either relegat

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

Bridging Classification and Reconstruction: Cooperative Time Series Anomaly Detection

DGX agent

arXiv:2605.26193v1 Announce Type: cross Abstract: Time series anomaly detection (TSAD) has long been a hot research topic in data mining due to its various applications. Recent studies challenge the e

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Cesarean Scar Defect Segmentation in Transvaginal Ultrasound Images: a Dataset and Benchmark

DGX agent

arXiv:2605.26774v1 Announce Type: new Abstract: Cesarean Scar Defect (CSD) is one of the most prevalent complications following cesarean delivery. Transvaginal ultrasonography is widely used for prima

model-releasesarxiv-cs-cv
27 May 2026
Safety

CFG-OEC: Classifier Free Guidance with Orthogonal Error Correction

DGX agent

arXiv:2511.14075v2 Announce Type: replace-cross Abstract: Classifier free guidance is a standard method for conditional sampling in diffusion models, but its sampling rule is not aligned with the obje

safetyarxiv-cs-ai
27 May 2026
Tutorials

Chaos-SSL: An Attention-Based Self-Supervised Learning Framework with Chaotic Transformation for Medical Image Classification

DGX agent

arXiv:2605.27146v1 Announce Type: new Abstract: Self-Supervised Learning (SSL) has emerged as a powerful paradigm to mitigate the reliance on large, annotated datasets, a common bottleneck in medical

tutorialsarxiv-cs-cv
27 May 2026
Model Releases

CIRCLED: A Multi-turn CIR Dataset with Consistent Dialogues across Domains

DGX agent

arXiv:2605.26734v1 Announce Type: new Abstract: Existing Multi-Turn Composed Image Retrieval (MTCIR) datasets lack dialogue-history consistency and are restricted to the fashion domain. To address the

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

Clinically-Grounded Counterfactual Reasoning for Medical Video Diagnosis

DGX agent

arXiv:2605.26483v1 Announce Type: new Abstract: Medical video diagnosis involves inferring clinical decisions from dynamic tissue responses throughout examination processes. Existing methods rely on a

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

Cogent Security launches autonomous vulnerability response tools as AI-assisted exploits outpace scanners

DGX agent

Cogent Security Inc., a startup that employs agentic artificial intelligence for vulnerability management, today launched two new platform capabilities aimed at compressing enterprise vulnerability re

model-releasessiliconangle
27 May 2026
Research

Cordyceps: Covert Control Attacks on LLMs via Data Poisoning

DGX agent

arXiv:2605.26595v1 Announce Type: cross Abstract: Large language models (LLMs) are often fine-tuned on uncurated text datasets that adversaries can poison. Existing poisoning attacks primarily rely on

researcharxiv-cs-ai
27 May 2026
Safety

Counterfactual Credit Policy Optimization for Multi-Agent Collaboration

DGX agent

arXiv:2603.21563v2 Announce Type: replace Abstract: Collaborative multi-agent large language models (LLMs) can solve complex reasoning tasks by decomposing roles, but reinforcement learning for such s

safetyarxiv-cs-ai
27 May 2026
Model Releases

Deep-layer limit and stability analysis of the basic forward-backward-splitting induced network (II): learning problems

DGX agent

arXiv:2605.27133v1 Announce Type: cross Abstract: Deep unfolding neural networks derived from iterative optimization schemes and numerical ordinary/partial differential equations (ODEs/PDEs) have attr

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

DelowlightSplat: Feed-Forward Gaussian Splatting for Lowlight 3D Scene Reconstruction

DGX agent

arXiv:2605.26629v1 Announce Type: new Abstract: Novel-view synthesis and 3D reconstruction from sparse posed images are central to robotics and AR/VR. Yet, feed-forward 3D Gaussian reconstruction fail

model-releasesarxiv-cs-cv
27 May 2026
Tutorials

Dissecting Multimodal In-Context Learning: Modality Asymmetries and Circuit Dynamics in modern Transformers

DGX agent

arXiv:2601.20796v2 Announce Type: replace Abstract: Transformer-based multimodal large language models often exhibit in-context learning (ICL) abilities. Motivated by this phenomenon, we ask: how do t

tutorialsarxiv-cs-cl
27 May 2026
Model Releases

Distribution-Aware Conformal Prediction: A Framework for generating efficient prediction intervals for time series

DGX agent

arXiv:2605.26569v1 Announce Type: new Abstract: We present Distribution-aware Conformal Prediction (DCP), a unified framework integrating probabilistic predictors like Monte Carlo dropout, deep ensemb

model-releasesarxiv-cs-lg
27 May 2026
Model Releases

Doppel launches agentic email security to disrupt phishing campaigns at the source

DGX agent

Social engineering defense startup Doppel Inc. today launched Doppel Email Security, an agentic artificial intelligence layer that traces phishing messages back to attacker infrastructure and orchestr

model-releasessiliconangle
27 May 2026
Model Releases

E3: Issue-Level Backtesting for Automated Research Critique

DGX agent

arXiv:2605.27072v1 Announce Type: cross Abstract: We present E3, an automated review assistant that augments reviewers and engineering teams by identifying decision-relevant technical concerns in rese

model-releasesarxiv-cs-ai
27 May 2026
← Previous
1…837838839840841…1314
Next →