AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “engineering”

GridTimelineEvolution
5,412 results
Model Releases

Scaffold, Not Vocabulary? A Controlled, Two-Tier, Pre-Registered Study of a Popperian Code-Generation Skill

DGX agent

arXiv:2606.06454v1 Announce Type: cross Abstract: Large language models increasingly write, review, and judge code, and a fast-growing practice equips them with prompt 'skills' that ask the model to r

model-releasesarxiv-cs-cl
5 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

Shopify on Replit + the new SEO Agent https://x.com/i/broadcasts/1kJzDDopENZKv

DGX agent

This post likely covers a live broadcast or announcement discussing the integration of Shopify with Replit, along with information about a newly released SEO Agent tool. The content probably demonstra

agentsreplit--x
5 Jun 2026
Safety

The creator of Linux just publicly called out the AI hype. Word for word. Linus Torvalds took the stage at Open Source Summit 2026 and said …

DGX agent

The creator of Linux just publicly called out the AI hype. Word for word. Linus Torvalds took the stage at Open Source Summit 2026 and said this: 'When I see people saying 99% of our code is written b

safetygary-marcus--x
5 Jun 2026
Agents

// The Meta-Agent Challenge // How good are current agents at self-improving? This is a great paper covering some of the challenges. They pr…

DGX agent

// The Meta-Agent Challenge // How good are current agents at self-improving? This is a great paper covering some of the challenges. They propose the Meta-Agent Challenge (MAC), where they give a codi

agentsdair-ai--x
5 Jun 2026
Research

The Tell-Tale Norm: ell_2 Magnitude as a Signal for Reasoning Dynamics in Large Language Models

DGX agent

arXiv:2606.06188v1 Announce Type: new Abstract: Recent work has sought to understand Large Language Models (LLMs) reasoning, yet a principled, model-intrinsic signal that captures its layer-wise reaso

researcharxiv-cs-cl
5 Jun 2026
Applications

Today, we are officially launching the Sakana AI RSI Lab in Tokyo to build open-ended, adaptive AI systems that collectively self-improve. I…

DGX agent

Today, we are officially launching the Sakana AI RSI Lab in Tokyo to build open-ended, adaptive AI systems that collectively self-improve. I am incredibly proud of our team’s work over the past 2 year

applicationsdavid-ha--x
5 Jun 2026
Hardware

Towards Realistic 3D Sonar Simulation

DGX agent

arXiv:2606.06130v1 Announce Type: new Abstract: As underwater robotics research increasingly addresses complex 3D perception and autonomous navigation, the fidelity of sonar simulation has become a ke

hardwarearxiv-cs-ro
5 Jun 2026
Safety

UNIVID: Unified Vision-Language Model for Video Moderation

DGX agent

arXiv:2606.05748v1 Announce Type: cross Abstract: Global-scale video moderation faces a dual challenge: the need for fine-grained multi-modal reasoning and the demand for interpretable outputs to supp

safetyarxiv-cs-cl
5 Jun 2026
Local Ai

v0.30.6-rc0

DGX agent

v0.30.6-rc0 is a release candidate that fixes kernel template instantiation so library symbols are exported correctly , following improvements from the v0.30 series. The v0.30 base release improved co

local-aiollama-releases
5 Jun 2026
Model Releases

Video-Rate Streaming Stylization on a Vision-Aware MLLM-Conditioned Edit Diffusion: Asymmetric Batched Inference on a Distilled UNet + MLLM Text Encoder

DGX agent

arXiv:2606.05981v1 Announce Type: new Abstract: Aggressive distillation of the diffusion U-Net inverts the per-frame bottleneck of real-time text-to-image pipelines: once the denoiser is a 4-step or 1

model-releasesarxiv-cs-cv
5 Jun 2026
Safety

Anycast Performance in Context

DGX agent

arXiv:2606.04298v1 Announce Type: cross Abstract: IP anycast lets a service advertise one address from many physical sites, leaving BGP to map each client to a site. It is central to the DNS root serv

safetyarxiv-cs-ai
4 Jun 2026
Local Ai

b9515

DGX agent

llama.cpp is a C/C++ implementation designed to enable LLM inference with minimal setup and state-of-the-art performance on a wide range of hardware locally and in the cloud. Build b9515 is an interme

local-aillama-cpp-releases
4 Jun 2026
Agents

Beyond Correctness: Rewarding Faithful Reasoning in Retrieval-Augmented Generation

DGX agent

arXiv:2510.13272v3 Announce Type: replace Abstract: Inspired by the success of reinforcement learning (RL) in Large Language Model (LLM) training for domains like math and code, recent work has begun

agentsarxiv-cs-cl
4 Jun 2026
Applications

ClustRecNet: A Novel End-to-End Deep Learning Framework for Clustering Algorithm Recommendation

DGX agent

arXiv:2509.25289v4 Announce Type: replace-cross Abstract: Identifying an effective clustering algorithm for a given dataset remains a fundamental unsupervised learning issue. We introduce ClustRecNet,

applicationsarxiv-cs-ai
4 Jun 2026
Research

DeliChess: A Multi-party Dialogue Dataset for Deliberation in Chess Puzzle Solving

DGX agent

arXiv:2606.04987v1 Announce Type: cross Abstract: Multi-party dialogue is a critical setting for studying collaborative reasoning and decision-making, yet existing datasets rarely focus on structured,

researcharxiv-cs-ai
4 Jun 2026
Research

EpiFormer: Learning Antigen-Antibody Interactions for Epitope Prediction via Geometric Deep Learning

DGX agent

arXiv:2606.04154v1 Announce Type: cross Abstract: Antibodies neutralize foreign antigens by binding to specific surface regions called epitopes. Computational epitope prediction is critical for unders

researcharxiv-cs-lg
4 Jun 2026
Model Releases

Finally! the first eval ship from cog!!!!!!!!!! 👼🏼 To contextualize: @METR_Evals cap out at ~16 hours. Cog has private enterprise evals up…

DGX agent

Finally! the first eval ship from cog!!!!!!!!!! 👼🏼 To contextualize: @METR_Evals cap out at ~16 hours. Cog has private enterprise evals up to 100hrs, and is confident enough to put a financial guarant

model-releasesswyx--x
4 Jun 2026
Industry

Five takeaways from the Cisco Live keynotes

DGX agent

At Cisco Systems Inc.‘s annual event, Cisco Live, this week in Las Vegas, it was no surprise that artificial intelligence was the top theme of the show and dominated most of the news and product innov

industrysiliconangle
4 Jun 2026
Model Releases

From Symbolic to Geometric: Enabling Spatial Reasoning in Large Language Models

DGX agent

arXiv:2606.04381v1 Announce Type: cross Abstract: Recent large language models (LLMs) often appear to exhibit spatial reasoning ability; however, this capability is largely symbolic, arising from patt

model-releasesarxiv-cs-ai
4 Jun 2026
Tutorials

How Endava is redesigning software delivery around AI agents

DGX agent

Endava, a software services company, is leveraging AI agents to fundamentally transform its software delivery processes and workflows. The case study likely demonstrates how the company is implementin

tutorialsopenai
4 Jun 2026
Safety

I think @Levie is overstating the positive case for employment in the (near term) AI era but that most people have overstated the negative c…

DGX agent

I think @Levie is overstating the positive case for employment in the (near term) AI era but that most people have overstated the negative case, and that the truth is somewhere in between. Which is to

safetygary-marcus--x
4 Jun 2026
Safety

MorphoQuant: Modality-Aware Quantization for Omni-modal Large Language Models

DGX agent

arXiv:2606.04349v1 Announce Type: cross Abstract: Conventional Post-Training Quantization (PTQ) methods struggle with 4-bit Omni-modal Large Language Models (OLLMs) due to the extreme distribution het

safetyarxiv-cs-ai
4 Jun 2026
Model Releases

New Benchmarking Shows Limited Generalization Power of TCR Antigenic Epitope Prediction Models

DGX agent

arXiv:2606.04994v1 Announce Type: new Abstract: Accurate computational prediction of T cell receptor (TCR) antigen specificity would transform the study of T cell biology and enable scalable immune en

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

Not All Errors Are Equal: Consequence-Aware Reasoning Compute Allocation

DGX agent

arXiv:2606.04402v1 Announce Type: new Abstract: Modern reasoning models can allocate different amounts of test-time computation, such as thinking tokens, model calls, or compute budget, to different t

model-releasesarxiv-cs-ai
4 Jun 2026
Tutorials

ParetoPilot: Zero-Surrogate Offline Multi-Objective Optimization via Infer-Perturb-Guide Diffusion

DGX agent

arXiv:2606.04468v1 Announce Type: cross Abstract: Offline multi-objective optimization (Offline MOO) aims to discover novel Pareto-optimal designs based on static datasets without expensive environmen

tutorialsarxiv-cs-ai
4 Jun 2026
Applications

reiterating: 'We're using the more expensive models to explore. Once we scale some of these experiences, we'll look to bring in more efficie…

DGX agent

reiterating: 'We're using the more expensive models to explore. Once we scale some of these experiences, we'll look to bring in more efficient models that are more efficient on a token basis or are op

applicationsclem-delangue--x
4 Jun 2026
Safety

Rethinking Continual Experience Internalization for Self-Evolving LLM Agents

DGX agent

arXiv:2606.04703v1 Announce Type: new Abstract: Experience internalization converts contextual experience from past interactions into reusable parametric capability, offering a promising path toward c

safetyarxiv-cs-cl
4 Jun 2026
Agents

SePO: Self-Evolving Prompt Agent for System Prompt Optimization

DGX agent

arXiv:2606.04465v1 Announce Type: cross Abstract: System prompt optimization improves agent behavior without modifying the underlying model, yielding human-readable, model-agnostic instructions. Exist

agentsarxiv-cs-ai
4 Jun 2026
Research

Simulate, Reason, Decide: Scientific Reasoning with LLMs for Simulation-Driven Decision Making

DGX agent

arXiv:2606.04505v1 Announce Type: new Abstract: Scientific simulators are increasingly being integrated into LLM-driven systems for high-stakes simulation-driven decision-making. However, existing fra

researcharxiv-cs-ai
4 Jun 2026
Safety

Smart Transportation Without Neurons -- Fair Metro Network Expansion with Tabular Reinforcement Learning

DGX agent

arXiv:2606.04167v1 Announce Type: cross Abstract: We tackle the Metro Network Expansion Problem (MNEP), a subset of the Transport Network Design Problem (TNDP), which focuses on expanding metro system

safetyarxiv-cs-ai
4 Jun 2026
Tutorials

SurvPFN: Towards Foundation Models for Survival Predictions

DGX agent

arXiv:2606.04564v1 Announce Type: new Abstract: Tabular foundation models (TFMs) have made rapid progress in standard classification and regression, but time-to-event survival prediction tasks have re

tutorialsarxiv-cs-lg
4 Jun 2026
Safety

The Accountability Horizon: An Impossibility Theorem for Governing Human-Agent Collectives

DGX agent

arXiv:2604.07778v2 Announce Type: replace Abstract: Existing accountability frameworks for AI systems, legal, ethical, and regulatory, rest on a shared assumption: for any consequential outcome, at le

safetyarxiv-cs-ai
4 Jun 2026
Model Releases

The Meta-Agent Challenge: Are Current Agents Capable of Autonomous Agent Development?

DGX agent

arXiv:2606.04455v1 Announce Type: new Abstract: Current AI benchmarks evaluate agents on task execution within human-designed workflows. These evaluations fundamentally fail to measure a critical next

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

The Saturation Trap and the Subjectivity of Intervention Timing: Why Affect-Based Triggers and LLM Judges Fail to Time Interventions on Autonomous Agents

DGX agent

arXiv:2606.04296v1 Announce Type: new Abstract: As autonomous AI agents move from conversational systems to long-horizon software execution, runtime safety layers that decide when to interrupt an agen

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Unpredictable Safety: Domain-Dependent Compliance and the Transparency Gap in Open-Weight LLMs

DGX agent

arXiv:2606.04035v1 Announce Type: cross Abstract: We present a systematic study of domain-dependent safety behavior in open-weight LLMs: 7 standardized experiments across 7 ethical domains, testing 5

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

VAMPS: Visual-Assisted Mathematical Problem Solving Benchmark

DGX agent

arXiv:2606.04244v1 Announce Type: new Abstract: Multimodal large language models are increasingly capable of complex reasoning, yet their performance often degrades when they must externalize a proble

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

95% token reduction. 30x faster execution. 90%+ task completion. Today at #MSBuild, we announced a major shift to move reasoning upstream: P…

DGX agent

95% token reduction. 30x faster execution. 90%+ task completion. Today at #MSBuild, we announced a major shift to move reasoning upstream: Pinecone Nexus now integrates directly with @Microsoft OneLak

model-releasespinecone--x
3 Jun 2026
Research

A cross-domain tropical species dataset with Chinese vernacular names and CITES source links

DGX agent

arXiv:2606.03156v1 Announce Type: new Abstract: We describe a versioned cross-domain dataset of 410,499 active tropical species (working snapshot 2026-04-20) spanning three applied subdomains -- tropi

researcharxiv-cs-cl
3 Jun 2026
Model Releases

Acceptance-Test-Driven Evaluation Protocols for Business-Centric LLM Systems

DGX agent

arXiv:2606.02755v1 Announce Type: cross Abstract: Large language model (LLM) applications are increasingly expected to satisfy deterministic institutional requirements while relying on probabilistic g

model-releasesarxiv-cs-ai
3 Jun 2026
Local Ai

b9487

DGX agent

The search results show general llama.cpp information and references to other recent builds (like b9484), but the specific details for b9487 were not clearly accessible. Based on the context from llam

local-aillama-cpp-releases
3 Jun 2026
Model Releases

Chatbots Output Meaningful (but Problematic) Language

DGX agent

arXiv:2606.02973v1 Announce Type: new Abstract: Are utterances by AI chatbots meaningful? Concretely, if a user asks, say, Anthropic's agent Claude, 'What is the capital of Spain?' and Claude answers,

model-releasesarxiv-cs-cl
3 Jun 2026
Safety

Discovering autonomous quantum error correction via deep reinforcement learning

DGX agent

arXiv:2511.12482v2 Announce Type: replace-cross Abstract: Quantum error correction is essential for fault-tolerant quantum computing. However, standard methods relying on active measurements may intro

safetyarxiv-cs-lg
3 Jun 2026
Safety

Easy-to-Use Shielding for Reinforcement Learning

DGX agent

arXiv:2606.03804v1 Announce Type: new Abstract: Safe exploration is a key challenge in Reinforcement Learning (RL) that aims to prevent agents from making harmful decisions while exploring their envir

safetyarxiv-cs-lg
3 Jun 2026
Agents

Economy of Minds: Emerging Multi-Agent Intelligence with Economic Interactions

DGX agent

arXiv:2606.02859v1 Announce Type: cross Abstract: How can a population of agents self-orchestrate and self-adapt into stronger collective intelligence without centralized control? Inspired by Friedric

agentsarxiv-cs-ai
3 Jun 2026
Safety

Enhancing Operational Safety via Agentic Dialogue Hazard Identification Analysis

DGX agent

arXiv:2606.03812v1 Announce Type: new Abstract: Operational safety in high-stakes domains such as industrial process control, autonomous, and safety-critical systems, demand reliable hazard identifica

safetyarxiv-cs-ai
3 Jun 2026
Research

Estimating Central, Peripheral, and Temporal Visual Contributions to Human Decision Making in Atari Games

DGX agent

arXiv:2604.04439v2 Announce Type: replace-cross Abstract: We study how different visual information sources contribute to human decision making in dynamic visual environments. Using Atari-HEAD, a larg

researcharxiv-cs-cv
3 Jun 2026
Safety

Fairness Definitions and Metrics in Deep Reinforcement Learning for Drug Discovery in Healthcare: A Rapid Evidence Review

DGX agent

arXiv:2606.02902v1 Announce Type: cross Abstract: Deep reinforcement learning (DRL) is increasingly applied to de novo molecular design, but choices in data, rewards, and evaluation can yield uneven p

safetyarxiv-cs-lg
3 Jun 2026
Model Releases

Gender-Dependent Diagnostic Substitution in LLM Medical Triage: Same Symptoms, Unequal Urgency

DGX agent

arXiv:2606.03641v1 Announce Type: new Abstract: We investigate whether large language models produce different medical triage recommendations for identical neurological symptoms when only the patient'

model-releasesarxiv-cs-ai
3 Jun 2026
← Previous
1…7980818283…113
Next →