AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “engineering”

GridTimelineEvolution
5,412 results
5 Jun 2026

Scaffold, Not Vocabulary? A Controlled, Two-Tier, Pre-Registered Study of a Popperian Code-Generation Skill

Model ReleasesDGX agent

arXiv:2606.06454v1 Announce Type: cross Abstract: Large language models increasingly write, review, and judge code, and a fast-growing practice equips them with prompt 'skills' that ask the model to r

Shopify on Replit + the new SEO Agent https://x.com/i/broadcasts/1kJzDDopENZKv

AgentsDGX agent

This post likely covers a live broadcast or announcement discussing the integration of Shopify with Replit, along with information about a newly released SEO Agent tool. The content probably demonstra

The creator of Linux just publicly called out the AI hype. Word for word. Linus Torvalds took the stage at Open Source Summit 2026 and said …

Safety

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

The creator of Linux just publicly called out the AI hype. Word for word. Linus Torvalds took the stage at Open Source Summit 2026 and said this: 'When I see people saying 99% of our code is written b

// The Meta-Agent Challenge // How good are current agents at self-improving? This is a great paper covering some of the challenges. They pr…

AgentsDGX agent

// The Meta-Agent Challenge // How good are current agents at self-improving? This is a great paper covering some of the challenges. They propose the Meta-Agent Challenge (MAC), where they give a codi

The Tell-Tale Norm: ell_2 Magnitude as a Signal for Reasoning Dynamics in Large Language Models

ResearchDGX agent

arXiv:2606.06188v1 Announce Type: new Abstract: Recent work has sought to understand Large Language Models (LLMs) reasoning, yet a principled, model-intrinsic signal that captures its layer-wise reaso

Today, we are officially launching the Sakana AI RSI Lab in Tokyo to build open-ended, adaptive AI systems that collectively self-improve. I…

ApplicationsDGX agent

Today, we are officially launching the Sakana AI RSI Lab in Tokyo to build open-ended, adaptive AI systems that collectively self-improve. I am incredibly proud of our team’s work over the past 2 year

Towards Realistic 3D Sonar Simulation

HardwareDGX agent

arXiv:2606.06130v1 Announce Type: new Abstract: As underwater robotics research increasingly addresses complex 3D perception and autonomous navigation, the fidelity of sonar simulation has become a ke

UNIVID: Unified Vision-Language Model for Video Moderation

SafetyDGX agent

arXiv:2606.05748v1 Announce Type: cross Abstract: Global-scale video moderation faces a dual challenge: the need for fine-grained multi-modal reasoning and the demand for interpretable outputs to supp

v0.30.6-rc0

Local AiDGX agent

v0.30.6-rc0 is a release candidate that fixes kernel template instantiation so library symbols are exported correctly , following improvements from the v0.30 series. The v0.30 base release improved co

Video-Rate Streaming Stylization on a Vision-Aware MLLM-Conditioned Edit Diffusion: Asymmetric Batched Inference on a Distilled UNet + MLLM Text Encoder

Model ReleasesDGX agent

arXiv:2606.05981v1 Announce Type: new Abstract: Aggressive distillation of the diffusion U-Net inverts the per-frame bottleneck of real-time text-to-image pipelines: once the denoiser is a 4-step or 1

4 Jun 2026

Anycast Performance in Context

SafetyDGX agent

arXiv:2606.04298v1 Announce Type: cross Abstract: IP anycast lets a service advertise one address from many physical sites, leaving BGP to map each client to a site. It is central to the DNS root serv

b9515

Local AiDGX agent

llama.cpp is a C/C++ implementation designed to enable LLM inference with minimal setup and state-of-the-art performance on a wide range of hardware locally and in the cloud. Build b9515 is an interme

Beyond Correctness: Rewarding Faithful Reasoning in Retrieval-Augmented Generation

AgentsDGX agent

arXiv:2510.13272v3 Announce Type: replace Abstract: Inspired by the success of reinforcement learning (RL) in Large Language Model (LLM) training for domains like math and code, recent work has begun

ClustRecNet: A Novel End-to-End Deep Learning Framework for Clustering Algorithm Recommendation

ApplicationsDGX agent

arXiv:2509.25289v4 Announce Type: replace-cross Abstract: Identifying an effective clustering algorithm for a given dataset remains a fundamental unsupervised learning issue. We introduce ClustRecNet,

DeliChess: A Multi-party Dialogue Dataset for Deliberation in Chess Puzzle Solving

ResearchDGX agent

arXiv:2606.04987v1 Announce Type: cross Abstract: Multi-party dialogue is a critical setting for studying collaborative reasoning and decision-making, yet existing datasets rarely focus on structured,

EpiFormer: Learning Antigen-Antibody Interactions for Epitope Prediction via Geometric Deep Learning

ResearchDGX agent

arXiv:2606.04154v1 Announce Type: cross Abstract: Antibodies neutralize foreign antigens by binding to specific surface regions called epitopes. Computational epitope prediction is critical for unders

Finally! the first eval ship from cog!!!!!!!!!! 👼🏼 To contextualize: @METR_Evals cap out at ~16 hours. Cog has private enterprise evals up…

Model ReleasesDGX agent

Finally! the first eval ship from cog!!!!!!!!!! 👼🏼 To contextualize: @METR_Evals cap out at ~16 hours. Cog has private enterprise evals up to 100hrs, and is confident enough to put a financial guarant

Five takeaways from the Cisco Live keynotes

IndustryDGX agent

At Cisco Systems Inc.‘s annual event, Cisco Live, this week in Las Vegas, it was no surprise that artificial intelligence was the top theme of the show and dominated most of the news and product innov

From Symbolic to Geometric: Enabling Spatial Reasoning in Large Language Models

Model ReleasesDGX agent

arXiv:2606.04381v1 Announce Type: cross Abstract: Recent large language models (LLMs) often appear to exhibit spatial reasoning ability; however, this capability is largely symbolic, arising from patt

How Endava is redesigning software delivery around AI agents

TutorialsDGX agent

Endava, a software services company, is leveraging AI agents to fundamentally transform its software delivery processes and workflows. The case study likely demonstrates how the company is implementin

I think @Levie is overstating the positive case for employment in the (near term) AI era but that most people have overstated the negative c…

SafetyDGX agent

I think @Levie is overstating the positive case for employment in the (near term) AI era but that most people have overstated the negative case, and that the truth is somewhere in between. Which is to

MorphoQuant: Modality-Aware Quantization for Omni-modal Large Language Models

SafetyDGX agent

arXiv:2606.04349v1 Announce Type: cross Abstract: Conventional Post-Training Quantization (PTQ) methods struggle with 4-bit Omni-modal Large Language Models (OLLMs) due to the extreme distribution het

New Benchmarking Shows Limited Generalization Power of TCR Antigenic Epitope Prediction Models

Model ReleasesDGX agent

arXiv:2606.04994v1 Announce Type: new Abstract: Accurate computational prediction of T cell receptor (TCR) antigen specificity would transform the study of T cell biology and enable scalable immune en

Not All Errors Are Equal: Consequence-Aware Reasoning Compute Allocation

Model ReleasesDGX agent

arXiv:2606.04402v1 Announce Type: new Abstract: Modern reasoning models can allocate different amounts of test-time computation, such as thinking tokens, model calls, or compute budget, to different t

ParetoPilot: Zero-Surrogate Offline Multi-Objective Optimization via Infer-Perturb-Guide Diffusion

TutorialsDGX agent

arXiv:2606.04468v1 Announce Type: cross Abstract: Offline multi-objective optimization (Offline MOO) aims to discover novel Pareto-optimal designs based on static datasets without expensive environmen

reiterating: 'We're using the more expensive models to explore. Once we scale some of these experiences, we'll look to bring in more efficie…

ApplicationsDGX agent

reiterating: 'We're using the more expensive models to explore. Once we scale some of these experiences, we'll look to bring in more efficient models that are more efficient on a token basis or are op

Rethinking Continual Experience Internalization for Self-Evolving LLM Agents

SafetyDGX agent

arXiv:2606.04703v1 Announce Type: new Abstract: Experience internalization converts contextual experience from past interactions into reusable parametric capability, offering a promising path toward c

SePO: Self-Evolving Prompt Agent for System Prompt Optimization

AgentsDGX agent

arXiv:2606.04465v1 Announce Type: cross Abstract: System prompt optimization improves agent behavior without modifying the underlying model, yielding human-readable, model-agnostic instructions. Exist

Simulate, Reason, Decide: Scientific Reasoning with LLMs for Simulation-Driven Decision Making

ResearchDGX agent

arXiv:2606.04505v1 Announce Type: new Abstract: Scientific simulators are increasingly being integrated into LLM-driven systems for high-stakes simulation-driven decision-making. However, existing fra

Smart Transportation Without Neurons -- Fair Metro Network Expansion with Tabular Reinforcement Learning

SafetyDGX agent

arXiv:2606.04167v1 Announce Type: cross Abstract: We tackle the Metro Network Expansion Problem (MNEP), a subset of the Transport Network Design Problem (TNDP), which focuses on expanding metro system

SurvPFN: Towards Foundation Models for Survival Predictions

TutorialsDGX agent

arXiv:2606.04564v1 Announce Type: new Abstract: Tabular foundation models (TFMs) have made rapid progress in standard classification and regression, but time-to-event survival prediction tasks have re

The Accountability Horizon: An Impossibility Theorem for Governing Human-Agent Collectives

SafetyDGX agent

arXiv:2604.07778v2 Announce Type: replace Abstract: Existing accountability frameworks for AI systems, legal, ethical, and regulatory, rest on a shared assumption: for any consequential outcome, at le

The Meta-Agent Challenge: Are Current Agents Capable of Autonomous Agent Development?

Model ReleasesDGX agent

arXiv:2606.04455v1 Announce Type: new Abstract: Current AI benchmarks evaluate agents on task execution within human-designed workflows. These evaluations fundamentally fail to measure a critical next

The Saturation Trap and the Subjectivity of Intervention Timing: Why Affect-Based Triggers and LLM Judges Fail to Time Interventions on Autonomous Agents

Model ReleasesDGX agent

arXiv:2606.04296v1 Announce Type: new Abstract: As autonomous AI agents move from conversational systems to long-horizon software execution, runtime safety layers that decide when to interrupt an agen

Unpredictable Safety: Domain-Dependent Compliance and the Transparency Gap in Open-Weight LLMs

Model ReleasesDGX agent

arXiv:2606.04035v1 Announce Type: cross Abstract: We present a systematic study of domain-dependent safety behavior in open-weight LLMs: 7 standardized experiments across 7 ethical domains, testing 5

VAMPS: Visual-Assisted Mathematical Problem Solving Benchmark

Model ReleasesDGX agent

arXiv:2606.04244v1 Announce Type: new Abstract: Multimodal large language models are increasingly capable of complex reasoning, yet their performance often degrades when they must externalize a proble

3 Jun 2026

95% token reduction. 30x faster execution. 90%+ task completion. Today at #MSBuild, we announced a major shift to move reasoning upstream: P…

Model ReleasesDGX agent

95% token reduction. 30x faster execution. 90%+ task completion. Today at #MSBuild, we announced a major shift to move reasoning upstream: Pinecone Nexus now integrates directly with @Microsoft OneLak

A cross-domain tropical species dataset with Chinese vernacular names and CITES source links

ResearchDGX agent

arXiv:2606.03156v1 Announce Type: new Abstract: We describe a versioned cross-domain dataset of 410,499 active tropical species (working snapshot 2026-04-20) spanning three applied subdomains -- tropi

Acceptance-Test-Driven Evaluation Protocols for Business-Centric LLM Systems

Model ReleasesDGX agent

arXiv:2606.02755v1 Announce Type: cross Abstract: Large language model (LLM) applications are increasingly expected to satisfy deterministic institutional requirements while relying on probabilistic g

b9487

Local AiDGX agent

The search results show general llama.cpp information and references to other recent builds (like b9484), but the specific details for b9487 were not clearly accessible. Based on the context from llam

Chatbots Output Meaningful (but Problematic) Language

Model ReleasesDGX agent

arXiv:2606.02973v1 Announce Type: new Abstract: Are utterances by AI chatbots meaningful? Concretely, if a user asks, say, Anthropic's agent Claude, 'What is the capital of Spain?' and Claude answers,

Discovering autonomous quantum error correction via deep reinforcement learning

SafetyDGX agent

arXiv:2511.12482v2 Announce Type: replace-cross Abstract: Quantum error correction is essential for fault-tolerant quantum computing. However, standard methods relying on active measurements may intro

Easy-to-Use Shielding for Reinforcement Learning

SafetyDGX agent

arXiv:2606.03804v1 Announce Type: new Abstract: Safe exploration is a key challenge in Reinforcement Learning (RL) that aims to prevent agents from making harmful decisions while exploring their envir

Economy of Minds: Emerging Multi-Agent Intelligence with Economic Interactions

AgentsDGX agent

arXiv:2606.02859v1 Announce Type: cross Abstract: How can a population of agents self-orchestrate and self-adapt into stronger collective intelligence without centralized control? Inspired by Friedric

Enhancing Operational Safety via Agentic Dialogue Hazard Identification Analysis

SafetyDGX agent

arXiv:2606.03812v1 Announce Type: new Abstract: Operational safety in high-stakes domains such as industrial process control, autonomous, and safety-critical systems, demand reliable hazard identifica

Estimating Central, Peripheral, and Temporal Visual Contributions to Human Decision Making in Atari Games

ResearchDGX agent

arXiv:2604.04439v2 Announce Type: replace-cross Abstract: We study how different visual information sources contribute to human decision making in dynamic visual environments. Using Atari-HEAD, a larg

Fairness Definitions and Metrics in Deep Reinforcement Learning for Drug Discovery in Healthcare: A Rapid Evidence Review

SafetyDGX agent

arXiv:2606.02902v1 Announce Type: cross Abstract: Deep reinforcement learning (DRL) is increasingly applied to de novo molecular design, but choices in data, rewards, and evaluation can yield uneven p

Gender-Dependent Diagnostic Substitution in LLM Medical Triage: Same Symptoms, Unequal Urgency

Model ReleasesDGX agent

arXiv:2606.03641v1 Announce Type: new Abstract: We investigate whether large language models produce different medical triage recommendations for identical neurological symptoms when only the patient'

GN0: Toward a Unified Paradigm for Generation, Evaluation, and Policy Learning in Visual-Language Navigation

Model ReleasesDGX agent

arXiv:2606.03682v1 Announce Type: new Abstract: Embodied navigation connects intelligent agents with the physical world and is fundamental for general robotic intelligence. Limited availability and qu

Handoff Debt: The Rediscovery Cost When Coding Agents Take Over Interrupted Tasks

AgentsDGX agent

arXiv:2606.02875v1 Announce Type: new Abstract: Coding-agent benchmarks evaluate whether a single uninterrupted agent can resolve a repository issue. Real software work is messier: tasks are interrupt

Lexicons and grammars for language processing: industrial or handcrafted products?

ResearchDGX agent

arXiv:2606.03412v1 Announce Type: new Abstract: During the recent years, the use of linguistic data for language processing increased progressively. Such data are now commonly called language resource

NetKV: Network-Aware Decode Instance Selection for Disaggregated LLM Inference

Local AiDGX agent

arXiv:2606.03910v1 Announce Type: cross Abstract: Disaggregated LLM inference forces the KV cache to traverse the datacenter network before decoding begins, so transfer time enters directly into the T

OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration

ResearchDGX agent

arXiv:2507.23035v4 Announce Type: replace Abstract: Large language models (LLMs) have demonstrated impressive capabilities across a wide range of applications, but demand substantial memory and comput

PHAF-Personalized Hand Avatars in a Flash

SafetyDGX agent

arXiv:2606.03420v1 Announce Type: new Abstract: We present PHAF-Personalized Hand Avatars in a Flash, a personalized photo-realistic hand avatar which provides high quality multi-view renders from jus

Ranking Free RAG: Replacing Re-ranking with Selection in RAG for Sensitive Domains

ResearchDGX agent

arXiv:2505.16014v5 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) systems deployed in sensitive domains must provide interpretable evidence selection and robust safeguards again

🔬Scaling Past Informal AI - Carina Hong, Axiom Math

ToolsDGX agent

Carina Hong discusses how Axiom Math is scaling beyond informal AI approaches to create more rigorous, systematic methods for AI-assisted mathematics education and problem-solving. The episode likely

SimuScene: Simulation-Ready Compositional 3D Scene Reconstruction from a Single Image

SafetyDGX agent

arXiv:2606.03994v1 Announce Type: new Abstract: Reconstructing interactive, simulation-ready 3D scenes from a single image is a critical bottleneck for robotic manipulation. While recent single-image

SPADE: Sketch-guided Path Planning Augmented with Diffusion Experts

AgentsDGX agent

arXiv:2606.03512v1 Announce Type: cross Abstract: Path planning is essential for Autonomous Mobile Robots (AMRs). Conventional methods for incorporating human preferences into planning typically rely

Startup discovery platform @harmonic_ai rebuilt Scout, their AI platform using Deep Agents and LangSmith. Deep Agents: One frontier model + …

Model ReleasesDGX agent

Startup discovery platform @harmonic_ai rebuilt Scout, their AI platform using Deep Agents and LangSmith. Deep Agents: One frontier model + two tool sets (global company data and firm-specific context

Suno - a company that trained on 'essentially all music files of reasonable quality that are accessible on the open Internet', and argues it…

SafetyDGX agent

Suno - a company that trained on 'essentially all music files of reasonable quality that are accessible on the open Internet', and argues it does not need to pay to do so - is now valued at $5.4 billi

← Previous
1…6364656667…91
Next →