AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,608 results
29 May 2026

Here's everything you need to know about Replit in 60 seconds ⭐️ → Plain English prompts turned into real working software → End-to-end work…

ToolsDGX agent

Here's everything you need to know about Replit in 60 seconds ⭐️ → Plain English prompts turned into real working software → End-to-end workflow from UI to deployment → Real-time team collaboration wi

Jailbreaking and Mitigation of Vulnerabilities in Large Language Models

SafetyDGX agent

arXiv:2410.15236v4 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have transformed artificial intelligence by advancing natural language understanding and generation, enabling app

MEMENTO: Leveraging Web as a Learning Signal for Low-Data Domains

TutorialsDGX agent

arXiv:2605.29795v1 Announce Type: new Abstract: Real-world tasks often lack large labeled datasets, motivating extensive work on learning in low-data regimes. However, existing approaches such as few-

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Minimal Prompt Perturbations Lead to Code Vulnerabilities: Prompt Fragility and Hidden-State Signals in Coding LLMs

Model ReleasesDGX agent

arXiv:2605.29737v1 Announce Type: cross Abstract: LLM-based coding assistants are seeing rapid adoption, offering substantial gains in developer productivity. As organizations increasingly ship code t

New LFM2.5 8b A1b model!!

Local AiDGX agent

Liquid AI released LFM2.5-8B-A1B, a device-optimized model designed to power real-life applications on phones, laptops, PCs, robots, and lightweight server-side use-cases. The model is a fast, memory-

OptSkills: Learning Generalizable Optimization Skills from Problem Archetypes via Cluster-Based Distillation

Model ReleasesDGX agent

arXiv:2605.29829v1 Announce Type: new Abstract: Leveraging Large Language Models (LLMs) to automatically formulate and solve optimization problems from natural language has emerged as an efficient par

PEARL: Training Socratic Tutors with Pedagogically Aligned Reinforcement Learning

SafetyDGX agent

arXiv:2605.29582v1 Announce Type: cross Abstract: Large Language Models (LLMs) have shown promise as educational tutors, yet effective tutoring requires more than solving problems: it must provide pro

REST3D: Reconstructing Physically Stable 3D Scenes from a Single Image

SafetyDGX agent

arXiv:2605.30338v1 Announce Type: new Abstract: Reconstructing physically stable 3D scenes from a single RGB image enables casual images to be converted into simulation-ready digital assets for applic

Run Step 3.7 Flash on NVIDIA GPUs with Enterprise-Ready Multimodal AI

HardwareDGX agent

Step 3.7 Flash is a 198B-parameter Mixture-of-Experts vision-language model designed for enterprise-scale production workloads, featuring native image and video input, a 256k context window, and confi

SalsaAgent: A multimodal embodied language model for interactive dance generation

ResearchDGX agent

arXiv:2605.29219v1 Announce Type: new Abstract: Interaction between humanoids involves bidirectional and nonverbal reactivity, coordination and synchrony. Toward socially aware robots and interactive

The improvements run wide. Across all major European languages, Command A+ consistently pulls ahead of competitors on WMT24++ (xCOMET-XL): …

Model ReleasesDGX agent

The improvements run wide. Across all major European languages, Command A+ consistently pulls ahead of competitors on WMT24++ (xCOMET-XL): 🇫🇷 +2.4 pts in French 🇪🇸 +1.9 pts in Spanish 🇩🇪 +0.9 pts in G

The Price Reversal Phenomenon: When Cheaper Reasoning Models Cost More

Model ReleasesDGX agent

arXiv:2603.23971v2 Announce Type: replace-cross Abstract: Developers and consumers increasingly choose reasoning models (RMs) based on their listed API prices. However, how accurately do these prices

When LLM Reward Design Fails: Diagnostic-Driven Refinement for Sparse Structured RL

Model ReleasesDGX agent

arXiv:2605.28918v1 Announce Type: new Abstract: For sparse, structured reinforcement-learning tasks with semantic reward-function interfaces, LLM-generated reward shaping is better framed as debugging

yes, absolutely, many companies are experimenting. but also: most of those experiments are failing to yield significant RoI. (weird for an e…

SafetyDGX agent

yes, absolutely, many companies are experimenting. but also: most of those experiments are failing to yield significant RoI. (weird for an economist to not even ask or address that question.) Looks li

28 May 2026

AI researchers ran 15-day simulations of worlds governed by different AI models: Claude Sonnet 4.6 recorded no crimes, while Gemini 3 Flash had the most at 683 (Jake Angelo/Fortune)

Model ReleasesDGX agent

Jake Angelo / Fortune: AI researchers ran 15-day simulations of worlds governed by different AI models: Claude Sonnet 4.6 recorded no crimes, while Gemini 3 Flash had the most at 683 — Imagine a world

Are We Truly Innovating? A Qualitative and Quantitative Study of Originality in AI Research Papers

ResearchDGX agent

arXiv:2602.06054v3 Announce Type: replace Abstract: Assessing originality in AI research is arguably the most consequential yet least reliable step in peer review. Reviewer judgments of originality re

Claude Opus 4.8: 'a modest but tangible improvement'

Model ReleasesDGX agent

Anthropic shipped Claude Opus 4.8 today. My favourite thing about it is this note in the release announcement: Users will find Opus 4.8 to be a modest but tangible improvement on its predecessor. Ther

Codex Mobile is free for everyone now and Goal Mode left beta. What it means if you do not code

IndustryDGX agent

Codex Mobile is now available in preview on iOS and Android across all ChatGPT plans, including Free and Go, in all supported regions. The feature allows developers to follow and steer Codex while it

Commit to the Bit: Reactive Reinforcement Learning Done Right

SafetyDGX agent

arXiv:2605.28276v1 Announce Type: new Abstract: Reinforcement learning algorithms are commonly analyzed (and designed) under the Markov assumption. This is unrealistic, as most environments encountere

Data Formulator 0.7: AI-powered data analytics for enterprise data

ApplicationsDGX agent

Data Formulator introduces AI-powered analytics for enterprise data workflows. Data teams can easily bring enterprise data into an AI-ready workspace where users can explore, analyze, and visualize da

Debate with Images: Detecting Deceptive Behaviors in Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2512.00349v2 Announce Type: replace Abstract: Are frontier AI systems becoming more capable? Certainly. Yet such progress is not an unalloyed blessing but rather a Trojan horse: behind their per

Delay-Aware Reinforcement Learning for Highway On-Ramp Merging under Stochastic Communication Latency

SafetyDGX agent

arXiv:2403.11852v5 Announce Type: replace-cross Abstract: Delayed and partially observable state information poses significant challenges for reinforcement learning (RL)-based control in real-world au

DisasterBench: Benchmarking LLM Planning under Typed Tool Interface Constraints

Model ReleasesDGX agent

arXiv:2605.27957v1 Announce Type: new Abstract: Disasters cause severe societal impacts, demanding rapid coordination of heterogeneous AI tools, from satellite analysis to flood prediction and damage

EchoAvatar: Real-time Generative Avatar Animation from Audio Streams

ResearchDGX agent

arXiv:2605.28272v1 Announce Type: new Abstract: Real-time synthesis of high-fidelity 3D character motion from audio is a pivotal component for next-generation interactive avatars and virtual assistant

ESC-Skills: Discovering and Self-Evolving Skills for Emotional Support Conversations

Model ReleasesDGX agent

arXiv:2605.27908v1 Announce Type: cross Abstract: Existing emotional support conversation (ESC) systems mainly rely on end-to-end response generation or coarse strategy supervision, offering limited i

Evolving Dataflow to process massive datasets for machine learning

Model ReleasesDGX agent

Google created MapReduce more than 20 years ago to solve the scaling problems in data processing that the then young company was running into. The AI era that we are in now demands efficient, large-sc

How Far Can Disaggregation Go? A Design-Space Exploration of Attention-FFN Disaggregation for Efficient MoE LLM Serving

Model ReleasesDGX agent

arXiv:2605.28302v1 Announce Type: cross Abstract: Modern large language model (LLM) inference has progressively disaggregated to keep pace with growing model sizes and tight TTFT and TPOT service-leve

just noticed today - the dataset is already past 1k+ downloads. opensource / openresearch ftw ! @evo__hq would be opensourcing as many datas…

Model ReleasesDGX agent

just noticed today - the dataset is already past 1k+ downloads. opensource / openresearch ftw ! @evo__hq would be opensourcing as many datasets, evals and autoresearch runs as we can in our pursuit of

Local Reachy Mini conversations wireless looks like magic! You can bring your friend around the house and get WOW effect from anyone! 🔥 Tha…

Model ReleasesDGX agent

Local Reachy Mini conversations wireless looks like magic! You can bring your friend around the house and get WOW effect from anyone! 🔥 Thanks @andimarafioti for the blog post on how to set this up: h

OmniVerifier-M1: Multimodal Meta-Verifier with Explicit Structured Recalibration

Local AiDGX agent

arXiv:2605.28805v1 Announce Type: cross Abstract: Visual outcomes are increasingly central to multimodal large language models, making reliable and fine-grained verification essential for scaling gene

OphIn-500K: Curating Web-Scale Visual Instructions for Scaling Ophthalmic Multimodal Large Language Models

ApplicationsDGX agent

arXiv:2605.27916v1 Announce Type: cross Abstract: The advancement of general medical Multimodal Large Language Models (MLLMs) has shown great potential for building conversational assistants to suppor

OralAgent: Integrating Reasoning, Tools, and Knowledge for Interactive Dental Image Analysis

Model ReleasesDGX agent

arXiv:2605.27378v1 Announce Type: new Abstract: Dental image analysis plays a pivotal role in supporting accurate diagnosis and treatment planning in oral healthcare. Although recent advances have pro

Pittsburgh-based Gray Swan, which stress-tests AI models for top frontier AI labs, raised a 40M Series A at a 200M valuation co-led by Wing VC and Madrona (Rashi Shrivastava/Forbes)

Model ReleasesDGX agent

Rashi Shrivastava / Forbes: Pittsburgh-based Gray Swan, which stress-tests AI models for top frontier AI labs, raised a 40M Series A at a 200M valuation co-led by Wing VC and Madrona — Gray Swan works

Plug-and-Play Benchmarking of Reinforcement Learning Algorithms for Large-Scale Flow Control

Model ReleasesDGX agent

arXiv:2601.15015v2 Announce Type: replace Abstract: Reinforcement learning (RL) has shown promising results in active flow control (AFC), yet progress in the field remains difficult to assess as exist

POINav: Benchmarking and Enhancing Final-Meters Arrival in Real-World Vision-Language Navigation

Model ReleasesDGX agent

arXiv:2605.28237v1 Announce Type: cross Abstract: Real-world navigation is fundamentally driven by Points of Interest (POIs), yet reaching a precise POI remains a critical 'final-meters' challenge. Ex

Snowveil: A Framework for Decentralised Preference Discovery

Model ReleasesDGX agent

arXiv:2512.18444v2 Announce Type: replace-cross Abstract: Aggregating subjective preferences in social choice traditionally assumes a trusted central authority. In contrast, this paper formalises Dece

Teacher-Student Representational Alignment for Reinforcement Learning-Driven Imitation Learning

SafetyDGX agent

arXiv:2605.28372v1 Announce Type: new Abstract: Imitation learning (IL) from a state-based reinforcement learning (RL) policy is a common approach to overcome the curse of dimensionality in complex an

The Illusion of Opting in AI-Mediated Consequential Decisions

SafetyDGX agent

arXiv:2605.28210v1 Announce Type: new Abstract: Drawing on Ullmann-Margalit's concept of opting (transformative, irrevocable, and shadowed by foreclosed alternatives), we show that current AI systems

“The mechanism is always the same in every story I've been covering. The demo works in a controlled environment with clean inputs. The deplo…

TutorialsDGX agent

“The mechanism is always the same in every story I've been covering. The demo works in a controlled environment with clean inputs. The deployment fails because real kitchens, real intersections, and r

VLA-Hijack: A Transferable Patch Attack against Vision-Language-Action Models via Visual Proprioception Hijacking

SafetyDGX agent

arXiv:2605.28083v1 Announce Type: new Abstract: While Vision-Language-Action (VLA) models have emerged as powerful generalist policies, their severe vulnerability to adversarial patches significantly

27 May 2026

1/ We’ve raised over 1B at a 26B valuation, led by @Lux_Capital, @generalcatalyst, and @8vc. Our enterprise usage has grown >10x since the…

ApplicationsDGX agent

1/ We’ve raised over 1B at a 26B valuation, led by @Lux_Capital, @generalcatalyst, and @8vc. Our enterprise usage has grown >10x since the start of this year, and our run-rate revenue grew to $492 M.

Alignment Tuning for Large Language Models: A Data-Centric Lens on Alignment Data Pipelines

SafetyDGX agent

arXiv:2605.26442v1 Announce Type: cross Abstract: Much of the alignment tuning literature is organized around optimization objectives, while the construction of alignment data is often treated implici

AssetGen: Deployable 3D Asset Generation at Interactive Speed

HardwareDGX agent

arXiv:2605.26137v1 Announce Type: cross Abstract: While 3D generation is progressing rapidly, recent work has often focused on obtaining high-resolution assets, leaving user experience and deployabili

Breaking the Epistemic Trap: Active Perception Under Compound Uncertainty

SafetyDGX agent

arXiv:2605.26627v1 Announce Type: cross Abstract: Deploying reinforcement learning in safety critical domains, from autonomous vehicles to medical decision support, is constrained by failures arising

ChartAct: A Benchmark for Dynamic Chart Understanding

Model ReleasesDGX agent

arXiv:2605.26994v1 Announce Type: new Abstract: Charts are widely used to present complex data for analysis and decision making. Existing chart understanding benchmarks mainly focus on static charts,

EgoProx: Evaluating MLLMs on Egocentric 3D Proximity Reasoning Across a Cognitive Hierarchy

Model ReleasesDGX agent

arXiv:2605.24456v2 Announce Type: replace Abstract: Humans constantly reason about 3D proximity, the relations between their body and surrounding objects, to guide perception and action in daily life.

Examining the Challenges of Intellectual Property in AI-Generated Productions

ApplicationsDGX agent

arXiv:2605.26590v1 Announce Type: cross Abstract: With the advancement of artificial intelligence systems capable of autonomously generating artistic, literary, musical works, and even inventions with

Exclusive: Unravel Data launches autonomous optimization engine for Databricks, Snowflake and BigQuery

Model ReleasesDGX agent

Unravel Data Systems Inc. is expanding beyond observability and FinOps software with a new autonomous optimization engine designed to automatically tune and remediate enterprise data platforms running

Explainable Cross-Disease Reasoning for Cardiovascular Risk Assessment from Low-Dose Computed Tomography

Local AiDGX agent

arXiv:2511.06625v5 Announce Type: replace-cross Abstract: Low-dose chest computed tomography (LDCT) captures pulmonary and cardiac structures in a single scan, enabling joint assessment of lung and ca

How AWS SMGS uses an AI-powered conversational assistant to transform business management with Amazon Bedrock AgentCore

TutorialsDGX agent

In this post, we share how we built NarrateAI using Amazon Bedrock AgentCore to deliver business intelligence at scale for the AWS SMGS (Sales, Marketing and Global Services) organization. You will le

'If I had Cofounder three years ago,' says @yoheinakajima, 'BabyAGI might have been a company.' Our newest case study tracks our most active…

ApplicationsDGX agent

'If I had Cofounder three years ago,' says @yoheinakajima, 'BabyAGI might have been a company.' Our newest case study tracks our most active user, Yohei Nakajima, building @ActiveGraphAI - the event s

if you’re working on a side project that you might want to incorporate later, I highly suggest you start using http://cofounder.co

ApplicationsDGX agent

if you’re working on a side project that you might want to incorporate later, I highly suggest you start using http://cofounder.co 'If I had Cofounder three years ago,' says @yoheinakajima, 'BabyAGI m

Innovative Silicosis and Pneumonia Classification: Leveraging Graph Transformer Post-hoc Modeling and Ensemble Techniques

ResearchDGX agent

arXiv:2501.00520v2 Announce Type: replace Abstract: This paper presents a comprehensive study on the classification and detection of Silicosis-related lung inflammation. Our main contributions include

Introducing Runway MCP. Now you can connect Runway directly into Claude, ChatGPT, Cursor, Replit and more. Generate polished images and vide…

Model ReleasesDGX agent

Introducing Runway MCP. Now you can connect Runway directly into Claude, ChatGPT, Cursor, Replit and more. Generate polished images and videos with state-of-the-art models, like Gen-4.5, Seedance 2.0,

Learning to Predict Future-Aligned Research Proposals with Language Models

Model ReleasesDGX agent

arXiv:2603.27146v3 Announce Type: replace Abstract: Large language models (LLMs) are increasingly used to assist ideation in research, but evaluating the quality of LLM-generated research proposals re

LLM-guided Hierarchical Search for End-to-end Reasoning Intensive Retrieval

Model ReleasesDGX agent

arXiv:2510.13217v2 Announce Type: replace-cross Abstract: Search systems are increasingly used for reasoning-intensive queries, where what makes a document relevant requires understanding or reasoning

LURE: Live-Usage Replay Evaluations for Reducing Evaluation Awareness

Model ReleasesDGX agent

arXiv:2605.26438v1 Announce Type: cross Abstract: Large language models can recognize when they are being evaluated (evaluation awareness) and behave differently because of that, which undermines the

Memory Architectures for Multi-Turn Text-to-SQL: A Benchmark and Empirical Study

Model ReleasesDGX agent

arXiv:2605.26394v1 Announce Type: new Abstract: Multi-turn Text-to-SQL is central to enterprise analytics yet remains predominantly evaluated in single-turn settings. We introduce EnterpriseMem-Bench,

Model-Harness-Task fit! it’s clear that RL post-training produces a model-harness fit via tool shapes and prompting as models are trained wi…

Model ReleasesDGX agent

Model-Harness-Task fit! it’s clear that RL post-training produces a model-harness fit via tool shapes and prompting as models are trained with the harness in the loop. Mentioned this in a previous Lan

ORCA: An End-to-End Interactive Copilot for Optimized Root Cause Analysis

TutorialsDGX agent

arXiv:2605.27022v1 Announce Type: new Abstract: Causal analysis is a crucial task in many domains, including manufacturing, social science, and medicine. However, despite recent progress, the conceptu

← Previous
1…274275276277278…294
Next →