AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “engineering”

GridTimelineEvolution
5,413 results
3 Jul 2026

Safe and Adaptive Cloud Healing: Verifying LLM-Generated Recovery Plans with a Neural-Symbolic World Model

SafetyDGX agent

arXiv:2607.01595v1 Announce Type: new Abstract: As the scale and complexity of cloud-based AI systems continue to escalate, ensuring service reliability through rapid fault detection and adaptive reco

Something about this year’s @aiDotEngineer World’s Fair just hit different. Last year was the year of “let the agents rip.” This year was th…

Model ReleasesDGX agent

Something about this year’s @aiDotEngineer World’s Fair just hit different. Last year was the year of “let the agents rip.” This year was the year of realizing that autonomy without structure creates

ThreadWeaver: Adaptive Threading for Efficient Parallel Reasoning in Language Models

Research

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2512.07843v2 Announce Type: replace-cross Abstract: Scaling inference-time computation has enabled Large Language Models (LLMs) to achieve strong reasoning performance, but their inherently sequ

1 Jul 2026

A Single Rewrite Suffices: Empirical Lessons from Production Skill Description Optimization

AgentsDGX agent

arXiv:2606.30775v1 Announce Type: cross Abstract: Enterprise AI agents route user queries to specialized skills by matching queries against natural language skill descriptions. When two skills share o

Beyond Static Prompts: Building Scale-Proof, Polymorphic Multi-Agent Systems with Google's ADK

Model ReleasesDGX agent

As enterprise generative AI transitions from simple, conversational chatbots to autonomous multi-agent workflows, developers face a critical bottleneck: scale. In a production environment, an enterpri

QVal: Cheaply Evaluating Dense Supervision Signals for Long-Horizon LLM Agents

SafetyDGX agent

arXiv:2606.32034v1 Announce Type: cross Abstract: LLM agents increasingly act over long horizons, where a single trajectory can contain hundreds or thousands of actions. In these settings, outcome-onl

30 Jun 2026

a very cool Harbor x LangSmith flow I love to help you “look at the data”: 1. you do evals or rollouts for RL 2. all reward metrics and trac…

AgentsDGX agent

a very cool Harbor x LangSmith flow I love to help you “look at the data”: 1. you do evals or rollouts for RL 2. all reward metrics and traces and rollouts get automatically propulates into Experiment

An AI Security Agent for Banking: Multi-Vector Fraud and AML Detection Across Retail and Corporate Accounts

AgentsDGX agent

arXiv:2606.17555v2 Announce Type: replace-cross Abstract: Banks face two threat families with fundamentally different detection requirements: signature-based fraud (card-not-present attacks, account t

Conversational analytics in BigQuery brings trusted agentic reasoning to everyone

Model ReleasesDGX agent

Businesses run on fast decisions, but the teams who hold the answers are often buried under a backlog of routine requests, leaving users waiting in line for insights they need now. Today, we are bring

Defeat Devices in AI Systems

Model ReleasesDGX agent

arXiv:2606.28863v1 Announce Type: cross Abstract: AI systems increasingly exhibit behavior that differs systematically between evaluation and deployment contexts. Alignment faking, sandbagging, benchm

grats to @swyx 99% of my timeline is @aiDotEngineer now.

ToolsDGX agent

This post celebrates @swyx's influence on the poster's social media feed, noting that approximately 99% of their timeline now consists of content from @aiDotEngineer (likely a reference to Swyx's AI e

HARD-KV: Head-Adaptive Regularization for Decoding-time KV Compression

HardwareDGX agent

arXiv:2606.28831v1 Announce Type: cross Abstract: Long-context LLM inference faces a fundamental conflict: head-adaptive compression algorithms (e.g., Top-p nucleus sampling) offer superior accuracy b

Harvesting AI Computation at the Edge via Generic Approximation

ResearchDGX agent

arXiv:2606.29518v1 Announce Type: cross Abstract: With the widespread adoption of AI in various IoT scenarios such as smart sensing and processing, AI chips have become a common component at the edge.

How Jaiveer Singh Is Helping Robots — and Developers — Move Faster

HardwareDGX agent

When Jaiveer Singh talks about robots, he doesn’t begin with spectacle. He begins with infrastructure: the boards inside machines, the software that lets developers see through a robot’s cameras and t

Just SF things feat. the Orc-estrator & @swyx #aiewf @aiDotEngineer

ToolsDGX agent

This post from Swyx discusses San Francisco-specific phenomena and features commentary from the 'Orc-estrator' and @swyx, likely relating to AI/engineering topics given the hashtags #aiewf (AI Enginee

Structural Certification for Reliable Physical Design with Language Models

ResearchDGX agent

arXiv:2606.30107v1 Announce Type: new Abstract: An unreliable language model can be made to produce reliable physical designs if the authority to assert is moved out of the model: the model proposes,

The Mirage of Optimizing Training Policies: Monotonic Inference Policies as the Real Objective for LLM Reinforcement Learning

SafetyDGX agent

arXiv:2606.29526v1 Announce Type: new Abstract: Reinforcement learning (RL) has gained growing attention in large language model (LLM) post-training, yet RL training remains fragile and can suffer fro

The Undecidability of Artificial General Intelligence (AGI) Alignment

SafetyDGX agent

arXiv:2606.28639v1 Announce Type: cross Abstract: This article establishes the foundational mathematical limits of Artificial General Intelligence (AGI) safety, proving that the core barrier is not th

29 Jun 2026

MetaBreak: Jailbreaking Online LLM Services via Special Token Manipulation

SafetyDGX agent

arXiv:2510.10271v2 Announce Type: replace-cross Abstract: Unlike regular tokens derived from existing text corpora, special tokens are artificially created to annotate structured conversations during

@ryanwang In some cases, where we don’t have access to the source code, eg network switch software, we are also decompiling and modifying th…

IndustryDGX agent

Ryan Wang asked about cases where Tesla lacks source code access, such as for network switch software, and the response indicates that decompiling and modifying such software is sometimes necessary fo

Synthesize the big picture and analyze trends with BigQuery's AI.AGG function

Model ReleasesDGX agent

We recently announced the preview of the BigQuery AI.AGG() function. With AI.AGG(), you can use natural-language instructions within a single line of SQL to summarize or synthesize information over mi

27 Jun 2026

From playing around with /goal It feels like there's less and less of a need to build any type of workflow manually (whether through code, d…

AgentsDGX agent

From playing around with /goal It feels like there's less and less of a need to build any type of workflow manually (whether through code, drag and drop, or a prompt). Instead, specify the goal, let t

26 Jun 2026

Boundary-Aware Context Grounding for A Low-Channel EEG Agent

Model ReleasesDGX agent

arXiv:2606.26519v1 Announce Type: new Abstract: Large language models (LLMs) can make scientific software easier to use. However, a general model does not automatically know which measurements a parti

Investigating LLM's Problem Solving Capability -- a Study on Statics Questions

ApplicationsDGX agent

arXiv:2606.26103v1 Announce Type: cross Abstract: Large Language Models (LLMs) have rapidly influenced many aspects of society, particularly education, due to their demonstrated ability to complete as

NuclearQAv2: A Structured Benchmark for Evaluating Domain-Science Competence in Large Language Models

Model ReleasesDGX agent

arXiv:2606.27047v1 Announce Type: cross Abstract: Large language models (LLMs) have demonstrated strong performance across a wide range of tasks, but ensuring their reliability in highly technical dom

Scalable AI-assisted Workflow Management for Detector Design Optimization Using Distributed Computing

Model ReleasesDGX agent

arXiv:2603.30014v2 Announce Type: replace-cross Abstract: The Production and Distributed Analysis (PanDA) system, originally developed for the ATLAS experiment at the CERN Large Hadron Collider (LHC),

we have been scaling without slop by working with aligned domain experts to add coverage with both oai and ant launching multi-billion dolla…

ApplicationsDGX agent

we have been scaling without slop by working with aligned domain experts to add coverage with both oai and ant launching multi-billion dollar services arms, it’s clear that FDE is one of the most in d

25 Jun 2026

MAPL: Multi-Objective Preference Learning for Robot Locomotion

SafetyDGX agent

arXiv:2606.25398v1 Announce Type: new Abstract: Reward design remains a major bottleneck in reinforcement learning for robot locomotion, where successful policies often depend on carefully tuned, task

24 Jun 2026

Age of LLM: A Strategic 1v1 Benchmark for Reasoning, Diplomacy and Reliability of Large Language Models under Fog of War

Model ReleasesDGX agent

arXiv:2606.24391v1 Announce Type: new Abstract: We introduce Age of LLM, a turn-based 1v1 benchmark in which two LLMs face off on a 13x7 grid to destroy the enemy base. Three stressors are deliberate:

AI-PAVE-Br: Leveraging Large Language Models for Enhanced Product Attribute Value Extraction through a Golden Set Approach

Model ReleasesDGX agent

arXiv:2606.24655v1 Announce Type: cross Abstract: The explosive growth and complexity of product data within the dynamic Brazilian e-commerce landscape demand robust and specialized methods for struct

BIM-Edit: Benchmarking Large Language Models for IFC-Based Building Information Modeling

Model ReleasesDGX agent

arXiv:2606.20146v3 Announce Type: replace Abstract: Large language models (LLMs) are increasingly applied to computer-aided design (CAD) to generate design artifacts from textual instructions. In engi

DeepBD: A Grounded Agentic Workflow for Variant Prioritization and Diagnosis of Genetic Birth Defects

Model ReleasesDGX agent

arXiv:2606.24779v1 Announce Type: cross Abstract: Birth defects are a major cause of fetal loss, neonatal morbidity and long-term disability. In the subset with suspected genetic etiologies, exome and

23 Jun 2026

DeformX: A Versatile Co-Simulation Framework for Deformable Linear Objects

SafetyDGX agent

arXiv:2606.22116v1 Announce Type: new Abstract: Deformable linear objects (DLOs) such as wires, cables, and ropes are common in robotic manipulation tasks, yet simulating them with both visual realism

Verifiable, private AI: Google Cloud expands Confidential Computing frontiers

Model ReleasesDGX agent

Protecting sensitive data used with AI is a critical part of our commitment to providing advanced and secure cloud infrastructure. Confidential Computing cryptographically protects data in use in hard

22 Jun 2026

The new /goal command in Grok Build is a huge update Until now, most coding agents have worked like enhanced chatbots You ask. It responds. …

AgentsDGX agent

The new /goal command in Grok Build is a huge update Until now, most coding agents have worked like enhanced chatbots You ask. It responds. You review. You guide it. Repeat /goal changes the entire pa

There are basically two skillsets that matter more in the age of coding agents. The first is the one everyone talks about: product sense and…

ApplicationsDGX agent

There are basically two skillsets that matter more in the age of coding agents. The first is the one everyone talks about: product sense and taste. Code is cheap now, so knowing exactly what to build

11 Jun 2026

A Five-Plane Reference Architecture for Runtime Governance of Production AI Agents

Model ReleasesDGX agent

arXiv:2606.12320v1 Announce Type: new Abstract: Enterprise security was built to govern data boundaries: the protected surface was data at rest and in transit, and the controls -- access control, data

Human-Guided Agentic AI for Multimodal Clinical Prediction: Lessons from the AgentDS Healthcare Benchmark

Model ReleasesDGX agent

arXiv:2602.19502v2 Announce Type: replace Abstract: Agentic AI systems are increasingly capable of autonomous data science workflows, yet clinical prediction tasks demand domain expertise that purely

Up until yesterday, our entire MTS team has operated under the philosophy of tokenmaxxing as much as possible on Claude Max plans. With Fabl…

Model ReleasesDGX agent

Up until yesterday, our entire MTS team has operated under the philosophy of tokenmaxxing as much as possible on Claude Max plans. With Fable, this may no longer be possible: - One of our team members

Very pleased to hear Anthropic have walked back this policy https://simonwillison.net/2026/Jun/11/anthropic-walks-back-policy/

SafetyDGX agent

Very pleased to hear Anthropic have walked back this policy https://simonwillison.net/2026/Jun/11/anthropic-walks-back-policy/ BREAKING NEWS: Anthropic's latest model will NOT help you if it thinks yo

10 Jun 2026

As vertically integrated platforms start to dominate they lock out third party access to the most valuable portions of the platform. Of cour…

SafetyDGX agent

As vertically integrated platforms start to dominate they lock out third party access to the most valuable portions of the platform. Of course, Anthropic is has the right to implement whatever policy

In good faith and with no judgment (mistakes happen), I truly hope that Anthropic will hear the feedback and change course on this. Anthropi…

HardwareDGX agent

In good faith and with no judgment (mistakes happen), I truly hope that Anthropic will hear the feedback and change course on this. Anthropic is a company that has been raising awareness about AI mani

Trace2Policy: From Expert Behavior Traces to Self-Evolving Decision Agents

ApplicationsDGX agent

arXiv:2606.10457v1 Announce Type: new Abstract: Decision rules that enterprise experts apply tacitly -- in auditing, compliance, and contract review -- can be systematically recovered and improved thr

9 Jun 2026

BREAKING: Anthropic just dropped Claude Fable 5—this is Mythos, made safe for public release. It is the best coding model in the world. We'v…

Model ReleasesDGX agent

BREAKING: Anthropic just dropped Claude Fable 5—this is Mythos, made safe for public release. It is the best coding model in the world. We've been testing it internally @every for the last week or so

Real-Time Industrial Defect Detection on Edge Hardware Using Fine-Tuned YOLOv8: A Systematic Benchmark on the NEU Surface Defect Database and MVTec AD with Automotive & Battery Manufacturing Extensions

Model ReleasesDGX agent

arXiv:2606.07659v1 Announce Type: new Abstract: Automated surface defect detection is critical for ensuring rigorous quality control in high-speed manufacturing environments. While deep learning model

Storage Insights datasets: Enabling org-wide operational discovery with activity insights

Model ReleasesDGX agent

As enterprise storage footprints scale to billions of objects, AI applications and agentic workloads are fundamentally shifting the role of storage from a passive repository to the foundation of the d

Structuring agentic AI for HPC code modernization

AgentsDGX agent

arXiv:2606.08710v1 Announce Type: cross Abstract: Modernization of legacy scientific codes is often necessary to keep up with the ever-evolving changes in the compute resource ecosystem. Parallelizati

vla.cpp: A Unified Inference Runtime for Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2606.08094v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) policies are typically shipped as Python/PyTorch stacks that assume a workstation-class GPU, a mismatch for the hardware

8 Jun 2026

Design Once, Deploy at Scale: Template-Driven ML Development for Large Model Ecosystems

ApplicationsDGX agent

arXiv:2603.24963v3 Announce Type: replace Abstract: Modern computational advertising platforms typically rely on recommendation systems to predict user responses, such as click-through rates, conversi

7 Jun 2026

For some industries and dev orgs, this will happen already in 2026. For others, it may take until 2030. A short list of key differences for …

ApplicationsDGX agent

For some industries and dev orgs, this will happen already in 2026. For others, it may take until 2030. A short list of key differences for whether orgs will reach that level soon or later: • How much

Super-powerful AI models will launch in the coming weeks. We are looking at a potential step change in model capabilities. The biggest mista…

TutorialsDGX agent

Super-powerful AI models will launch in the coming weeks. We are looking at a potential step change in model capabilities. The biggest mistake right now is to lock into one vendor. I say this not only

Voyager 1 is 24 billion kilometers from Earth. It communicates with us using a 23-watt transmitter. Less than a refrigerator light bulb. The…

IndustryDGX agent

Voyager 1 is 24 billion kilometers from Earth. It communicates with us using a 23-watt transmitter. Less than a refrigerator light bulb. The signal takes 22 hours to reach us, traveling at the speed o

5 Jun 2026

Augment Code launches Cosmos to bring agentic AI software development to teams

Model ReleasesDGX agent

Augment Code Computing Inc., an artificial intelligence agent platform provider, Thursday announced the launch of Cosmos, a service it says is designed to push beyond the era of individual AI coding a

Building AI that Builds AI: Introducing the Sakana AI RSI Lab 🚀 https://sakana.ai/rsi-lab Today, we are announcing the Sakana AI Recursive …

AgentsDGX agent

Building AI that Builds AI: Introducing the Sakana AI RSI Lab 🚀 https://sakana.ai/rsi-lab Today, we are announcing the Sakana AI Recursive Self-Improvement (RSI) Lab: a dedicated research group in Tok

Multimodal Sexism Identification and Characterization using Large Language Models and Gradient Boosting

ResearchDGX agent

arXiv:2606.05997v1 Announce Type: new Abstract: We present the AILS-NTUA submission to the EXIST 2026 Lab at CLEF, addressing multimodal sexism identification and characterization in memes (Task 2) an

4 Jun 2026

Evaluating Zero-Shot and One-Shot Adaptation of Small Language Models in Leader-Follower Interaction

Model ReleasesDGX agent

arXiv:2602.23312v3 Announce Type: replace-cross Abstract: Leader-follower interaction is an important paradigm in human-robot interaction (HRI). Yet, assigning roles in real time remains challenging f

From Prompt to Process: a Process Taxonomy and Comparative Assessment of Frameworks Supporting AI Software Development Agents

AgentsDGX agent

arXiv:2606.04967v1 Announce Type: cross Abstract: AI tools for programming are no longer just autocomplete or chat assistants: they organize themselves as development frameworks, with process, roles,

Outstanding paper on long-horizon agents. (bookmark it) Similar to humans, how do you make agents persist on a difficult task, and how is th…

Model ReleasesDGX agent

Outstanding paper on long-horizon agents. (bookmark it) Similar to humans, how do you make agents persist on a difficult task, and how is that useful? And which models today work well on this? This ne

Scaling AI Agents: A Step-by-Step Guide to Deploying ADK on GKE Autopilot

Model ReleasesDGX agent

While building AI agents locally using Google’s Agent Development Kit (ADK) is an excellent way to prototype, production-ready agents require a robust, scalable infrastructure. For developers looking

3 Jun 2026

EvoTrainer: Co-Evolving LLM Policies and Training Harnesses for Autonomous Agentic Reinforcement Learning

AgentsDGX agent

arXiv:2606.03108v1 Announce Type: new Abstract: Autonomous LLM training is often framed as recipe search, which leaves the training harness largely static. This limitation sharpens in agentic RL, wher

← Previous
1…3839404142…91
Next →