AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,376
  • Agents7,554
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,170
  • Local Ai4,930
  • Model Releases23,883
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,376
  • Agents7,554
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,170
  • Local Ai4,930
  • Model Releases23,883
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

88,376Total entries
1Added by human
88,375Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,602 results
11 May 2026

OpenAI just released its answer to Claude Mythos

Model ReleasesDGX agent

OpenAI is launching Daybreak, an AI initiative focused on detecting and patching vulnerabilities before attackers find them. Daybreak uses the Codex Security AI agent that launched in March to create

OrScale: Orthogonalised Optimization with Layer-Wise Trust-Ratio Scaling

Model ReleasesDGX agent

arXiv:2605.07815v1 Announce Type: cross Abstract: Muon improves neural-network training by orthogonalizing matrix-valued updates, but it leaves each layer's update magnitude controlled mostly by a glo

Path Integration and Object-Location Binding Emerge in an Action-Conditioned Predictive Sequence Network

TutorialsDGX agent

arXiv:2602.03490v2 Announce Type: replace Abstract: Adaptive cognition requires structured internal models of objects and their relations. Predictive neural networks are often proposed to learn such w

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

PRIMED: Adaptive Modality Suppression for Referring Audio-Visual Segmentation via Biased Competition

Model ReleasesDGX agent

arXiv:2605.07154v1 Announce Type: new Abstract: Referring Audio-Visual Segmentation (Ref-AVS) seeks to localize and segment target objects in video frames based on visual, auditory, and textual referr

Prototype Guided Post-pretraining for Single-Cell Representation Learning

TutorialsDGX agent

arXiv:2605.07938v1 Announce Type: new Abstract: Single-cell representation learning (SCRL) from gene expression data offers a way to uncover the complex regulatory logic underlying cellular function.

ProtSent: Protein Sentence Transformers

ResearchDGX agent

arXiv:2605.06830v1 Announce Type: cross Abstract: Protein language models (pLMs) produce per-residue representations that capture evolutionary and structural information, yet their mean-pooled sequenc

RCoT-Seg: Reinforced Chain-of-Thought for Video Reasoning and Segmentation

Model ReleasesDGX agent

arXiv:2605.07334v1 Announce Type: new Abstract: Video Reasoning Segmentation (VRS) aims to segment target objects in videos based on implicit instructions that convey human intent and temporal logic.

Reason to Play: Behavioral and Brain Alignment Between Frontier LRMs and Human Game Learners

SafetyDGX agent

arXiv:2605.08019v1 Announce Type: new Abstract: Humans rapidly learn abstract knowledge when encountering novel environments and flexibly deploy this knowledge to guide efficient and intelligent actio

Region4Web: Rethinking Observation Space Granularity for Web Agents

Model ReleasesDGX agent

arXiv:2605.07134v1 Announce Type: cross Abstract: Web agents perceive web pages through an observation space, yet its granularity has remained an underexamined design choice. Existing work treats obse

RL-RIG: A Generative Spatial Reasoner via Intrinsic Reflection

ResearchDGX agent

arXiv:2602.19974v2 Announce Type: replace Abstract: Recent advancements in image generation have achieved impressive results in producing high-quality images. However, existing image generation models

SatSurfGS: Generalizable 2D Gaussian Splatting for Sparse-View Satellite Surface Reconstruction

Model ReleasesDGX agent

arXiv:2605.07181v1 Announce Type: new Abstract: Sparse-view satellite image surface reconstruction remains highly challenging, fundamentally because the reliability of multi-view matching under satell

ScrapeGraphAI-100k: Dataset for Schema-Constrained LLM Generation

Model ReleasesDGX agent

arXiv:2602.15189v2 Announce Type: replace-cross Abstract: Producing output that conforms to a specified JSON schema underlies tool use, structured extraction, and knowledge base construction in modern

Search-based Robustness Testing of Laptop Refurbishing Robotic Software

Local AiDGX agent

arXiv:2605.07530v1 Announce Type: new Abstract: The Danish Technological Institute (DTI) focuses on transferring advanced technologies (including robots) to the industry and the public sector. One key

Semantic Integrity Matters: Benchmarking and Preserving High-Density Reasoning in KV Cache Compression

Model ReleasesDGX agent

arXiv:2502.01941v3 Announce Type: replace-cross Abstract: While Key-Value (KV) cache compression is essential for efficient LLM inference, current evaluations disproportionately focus on sparse retrie

SlopCodeBench: Benchmarking How Coding Agents Degrade Over Long-Horizon Iterative Tasks

Model ReleasesDGX agent

arXiv:2603.24755v2 Announce Type: replace-cross Abstract: Software development is iterative, yet agentic coding benchmarks hide design issues through their single-shot setup. Recent iterative benchmar

SREGym: A Live Benchmark for AI SRE Agents with High-Fidelity Failure Scenarios

Model ReleasesDGX agent

arXiv:2605.07161v1 Announce Type: new Abstract: AI agents are increasingly used to diagnose and mitigate failures in production systems, known as agentic Site Reliability Engineering (SRE). Current SR

SteelDefectX: A Multi-Form Vision-Language Dataset and Benchmark for Steel Surface Defect Analysis

Model ReleasesDGX agent

arXiv:2603.21824v2 Announce Type: replace-cross Abstract: Steel surface defect analysis is critical for industrial quality control, yet existing benchmarks rely primarily on label-only annotations, li

Structured Coupling for Flow Matching

TutorialsDGX agent

arXiv:2605.07676v1 Announce Type: new Abstract: Standard flow matching scales well but typically relies on an unstructured source distribution, limiting its ability to learn interpretable latent struc

The EDelta-MHC-Geo Transformer: Adaptive Geodesic Operations with Guaranteed Orthogonality

Model ReleasesDGX agent

arXiv:2605.06729v1 Announce Type: cross Abstract: We present the EDelta-MHC-Geo Transformer, a novel architecture that unifies Manifold-Constrained Hyper-Connections (mHC), Deep Delta Learning (DDL),

The Memory Curse: How Expanded Recall Erodes Cooperative Intent in LLM Agents

Model ReleasesDGX agent

arXiv:2605.08060v1 Announce Type: cross Abstract: Context window expansion is often treated as a straightforward capability upgrade for LLMs, but we find it systematically fails in multi-agent social

Topic Is Not Agenda: A Citation-Community Audit of Text Embeddings

Model ReleasesDGX agent

arXiv:2605.07158v1 Announce Type: cross Abstract: Vector search and retrieval-augmented generation (RAG) rest on the assumption that cosine similarity between text embeddings reflects conceptual relat

TraceFix: Repairing Agent Coordination Protocols with TLA+ Counterexamples

AgentsDGX agent

arXiv:2605.07935v1 Announce Type: new Abstract: We present TraceFix, a verification-first pipeline for Large Language Model (LLM) multi-agent coordination. An agent synthesizes a protocol topology as

Training-Induced Escape from Token Clustering in a Mean-Field Formulation of Transformers

Model ReleasesDGX agent

arXiv:2605.07772v1 Announce Type: new Abstract: Transformers perform inference by iteratively transforming token representations across layers. This layerwise computation has been studied empirically,

Using LLM in the shebang line of a script

Model ReleasesDGX agent

TIL: Using LLM in the shebang line of a script Kim_Bruning on Hacker News: But seriously, you can put a shebang on an english text file now (if you're sufficiently brave) [...] This inspired me to loo

Utility-Preserving De-Identification for Math Tutoring: Investigating Numeric Ambiguity in the MathEd-PII Benchmark Dataset

Model ReleasesDGX agent

arXiv:2602.16571v2 Announce Type: replace Abstract: Large-scale sharing of dialogue data is key to advancing the science of teaching and learning, yet rigorous de-identification remains a major barrie

Visual Text Compression as Measure Transport

Model ReleasesDGX agent

arXiv:2605.06708v1 Announce Type: cross Abstract: Visual text compression (VTC) promises efficient long-context processing by rendering text into an image and re-encoding it with a vision-language mod

📣We're calling for ambassadors! Whether you're a developer with great technical taste or a local community leader who loves bringing people…

Model ReleasesDGX agent

📣We're calling for ambassadors! Whether you're a developer with great technical taste or a local community leader who loves bringing people together, we'd love to have you join us. Visit the website b

When Losses Align: Gradient-Based Composite Loss Weighting for Efficient Pretraining

ResearchDGX agent

arXiv:2605.07756v1 Announce Type: cross Abstract: Modern deep models are often pretrained on large-scale data with missing labels using composite objectives, where the relative weights of multiple los

When Routine Chats Turn Toxic: Unintended Long-Term State Poisoning in Personalized Agents

Model ReleasesDGX agent

arXiv:2605.06731v1 Announce Type: cross Abstract: Personalized LLM agents maintain persistent cross-session state to support long-horizon collaboration. Yet, this persistence introduces a subtle but c

Zero-Shot Satellite Image Retrieval through Joint Embeddings: Application to Crisis Response

ResearchDGX agent

arXiv:2605.05405v2 Announce Type: replace Abstract: Semantic search of Earth observation archives remains challenging. Visual foundation models such as CLAY produce rich embeddings of satellite imager

10 May 2026

@roydanroy What he talks about couldn't have happened before GPT-5.5

Model ReleasesDGX agent

Sam Altman comments on a statement by @roydanroy, suggesting that the technological capabilities or developments being discussed would not have been possible prior to GPT-5.5. The post implies that GP

9 May 2026

Downloading now... 1M token context window with supposedly usable coding agent capability all on a 128GB Macbook Pro is 🤯

Model ReleasesDGX agent

Downloading now... 1M token context window with supposedly usable coding agent capability all on a 128GB Macbook Pro is 🤯 🚨 OPEN SOURCE AI IS LITERALLY UNSTOPPABLE 🚨 The legendary founder of Redis (An

Hot take on METR’s new graph that so many people are flipping about today. • Claude Code is a real advance; Mythos probably builds on some o…

Model ReleasesDGX agent

Hot take on METR’s new graph that so many people are flipping about today. • Claude Code is a real advance; Mythos probably builds on some of what is learned there. But… • If you read the graph carefu

This is confused, but popular. Popular because it tells a bunch of people what they want to hear. Confused for a couple reasons: first, Myth…

Model ReleasesDGX agent

This is confused, but popular. Popular because it tells a bunch of people what they want to hear. Confused for a couple reasons: first, Mythos probably isn’t a pure LLM. (Claude Code isn’t, and it pro

Yann LeCun closed $1.03B for AMI Labs on March 10. Three days later, this paper dropped from his NYU collaborators. 15M parameters. Single G…

HardwareDGX agent

Yann LeCun closed 1.03B for AMI Labs on March 10. Three days later, this paper dropped from his NYU collaborators. 15M parameters. Single GPU. A few hours of training. LeWorldModel is the first JEPA t

8 May 2026

Chain of thought monitors are a key layer of defense against AI agent misalignment. To preserve monitorability, we avoid penalizing misalign…

Model ReleasesDGX agent

Chain of thought monitors are a key layer of defense against AI agent misalignment. To preserve monitorability, we avoid penalizing misaligned reasoning during RL. We found a limited amount of acciden

I built an autonomous agent that lives inside her own source code— 7 days, 480 commits, multi-provider (DeepSeek / ChatGPT / Ollama)

Model ReleasesDGX agent

This post describes a project where the developer created an autonomous AI agent capable of modifying and executing its own source code across a 7-day development period, integrating multiple language

we'd like to help companies secure themselves and we think it's important to start work on this quickly

Model ReleasesDGX agent

we'd like to help companies secure themselves and we think it's important to start work on this quickly Today, we're rolling out GPT‑5.5‑Cyber in limited preview to defenders responsible for securing

7 May 2026

A Hybrid Quantum-Classical Framework for Financial Volatility Forecasting Based on Quantum Circuit Born Machines

TutorialsDGX agent

arXiv:2603.09789v2 Announce Type: replace Abstract: Accurate financial volatility forecasting is crucial but challenged by the non-linear, highly correlated nature of market data. Recently, quantum co

A Self-Attentive Meta-Optimizer with Group-Adaptive Learning Rates and Weight Decay

Model ReleasesDGX agent

arXiv:2605.04055v1 Announce Type: new Abstract: Adaptive optimizers like AdamW apply uniform hyperparameters across all parameter groups, ignoring heterogeneous optimization dynamics across layers and

A unified Benchmark for Multi-Frame Image Restoration under Severe Refractive Warping

Model ReleasesDGX agent

arXiv:2605.05079v1 Announce Type: new Abstract: Video sequence capturing through refractive dynamic media, such as a turbulent air or water surface, often suffer from severe geometric distortions and

Adapt to Thrive! Adaptive Power-Mean Policy Optimization for Improved LLM Reasoning

SafetyDGX agent

arXiv:2605.04066v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) is an essential paradigm that enhances the reasoning capabilities of Large Language Models (LLMs).

Adaptive Ensemble Aggregation for Actor-Critics

Model ReleasesDGX agent

arXiv:2507.23501v2 Announce Type: replace Abstract: Ensembles are ubiquitous in off-policy actor-critic learning, yet their efficacy depends critically on how they are aggregated. Current methods typi

Aggressive or Imperceptible, or Both: Network Pruning Assisted Hybrid Byzantines in Federated Learning

Model ReleasesDGX agent

arXiv:2404.06230v3 Announce Type: replace Abstract: In federated learning (FL), profiling and verifying each client is inherently difficult, which introduces a significant security vulnerability: mali

AnyPos: Automated Task-Agnostic Actions for Bimanual Manipulation

SafetyDGX agent

arXiv:2507.12768v2 Announce Type: replace Abstract: Learning generalizable manipulation policies hinges on data, yet robot manipulation data is scarce and often entangled with specific embodiments, ma

Automated Large-scale CVRP Solver Design via LLM-assisted Flexible MCTS

Model ReleasesDGX agent

arXiv:2605.03339v1 Announce Type: new Abstract: Solving large-scale CVRP (LSCVRP) with hundreds to thousands of nodes remains difficult for even state-of-the-art solvers. Divide-and-conquer can scale

Beyond Fixed Thresholds and Domain-Specific Benchmarks for Explainable Multi-Task Classification in Autonomous Vehicles

SafetyDGX agent

arXiv:2605.04299v1 Announce Type: new Abstract: Scene understanding is a vital part of autonomous driving systems, which requires the use of deep learning models. Deep learning methods are intrinsical

Beyond Public Access in LLM Pre-Training Data

SafetyDGX agent

arXiv:2505.00020v2 Announce Type: replace Abstract: Using a legally obtained dataset of 34 copyrighted O'Reilly Media books, we apply the DE-COP membership inference attack method to investigate wheth

CARD: A Multi-Modal Automotive Dataset for Dense 3D Reconstruction in Challenging Road Topography

Model ReleasesDGX agent

arXiv:2605.05014v1 Announce Type: new Abstract: Autonomous driving must operate across diverse surfaces to enable safe mobility. However, most driving datasets are captured on well-paved flat roads. M

Conceptors for Semantic Steering

Model ReleasesDGX agent

arXiv:2605.04980v1 Announce Type: cross Abstract: Activation-based steering provides control of LLM behavior at inference time, but the dominant paradigm reduces each concept to a single direction who

Confronting Label Indeterminacy in Automated Bail Decisions

SafetyDGX agent

arXiv:2605.04073v1 Announce Type: new Abstract: Bail decisions present a fundamental challenge for data-driven decision support systems. When bail is denied, the counterfactual outcome of whether the

Detecting Deepfakes via Hamiltonian Dynamics

ResearchDGX agent

arXiv:2605.04405v1 Announce Type: new Abstract: Driven by the rapid development of generative AI models, deepfake detectors are compelled to undergo periodic recalibration to capture newly developed s

Evaluating Prompting and Execution-Based Methods for Deterministic Computation in LLMs

ResearchDGX agent

arXiv:2605.03227v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated strong capabilities in natural language understanding and reasoning. However, their ability to perform ex

Few-Shot Learning Pipeline for Monkeypox Skin Disease Classification Using CNN Feature Extractors

Model ReleasesDGX agent

arXiv:2605.05034v1 Announce Type: new Abstract: Despite the strong performance of Convolutional Neural Networks (CNNs) in disease classification, their effectiveness often depends on access to large a

From SWE-ZERO to SWE-HERO: Execution-free to Execution-based Fine-tuning for Software Engineering Agents

Model ReleasesDGX agent

arXiv:2604.01496v2 Announce Type: replace-cross Abstract: We introduce SWE-ZERO to SWE-HERO, a two-stage SFT recipe that achieves state-of-the-art results on SWE-bench by distilling open-weight fronti

From Video-to-PDE: Data-Driven Discovery of Nonlinear Dye Plume Dynamics

ResearchDGX agent

arXiv:2605.04535v1 Announce Type: new Abstract: Inferring continuum models directly from video is hampered by two facts: the recorded field is uncalibrated image intensity rather than a physical state

Gaze4HRI: Zero-shot Benchmarking Gaze Estimation Neural-Networks for Human-Robot Interaction

Model ReleasesDGX agent

arXiv:2605.04770v1 Announce Type: new Abstract: While zero-shot appearance-based 3D gaze estimation offers significant cost-efficiency by directly mapping RGB images to gaze vectors, its reliability i

GEM: Graph-Enhanced Mixture-of-Experts with ReAct Agents for Dialogue State Tracking

AgentsDGX agent

arXiv:2605.04449v1 Announce Type: new Abstract: Dialogue State Tracking (DST) requires precise extraction of structured information from multi-domain conversations, a task where Large Language Models

Geometric Evolution Graph Convolutional Networks: Enhancing Graph Representation Learning via Ricci Flow

Model ReleasesDGX agent

arXiv:2603.26178v2 Announce Type: replace Abstract: We introduce the Geometric Evolution Graph Convolutional Network (GEGCN), a novel framework that enhances graph representation learning through expl

Geometry over Density: Few-Shot Cross-Domain OOD Detection

ApplicationsDGX agent

arXiv:2605.03410v2 Announce Type: new Abstract: Out-of-distribution (OOD) detection identifies test samples that fall outside a model's training distribution, a capability critical for safe deployment

← Previous
1…564565566567568…1061
Next →