AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “papers”

GridTimelineEvolution
12,157 results
11 Aug 2026

Private Etymology: Designing Relational Reuse of Shared Symbols in Long-Term Human-AI Interaction

Local AiDGX agent

arXiv:2608.08443v1 Announce Type: cross Abstract: Previous studies have shown that people can develop shared symbols, partner-specific expressions, personal idioms, inside jokes, and other parts of a

10 Aug 2026

Counterfactual Simulation Training for Chain-of-Thought Faithfulness

ResearchDGX agent

arXiv:2602.20710v2 Announce Type: replace Abstract: Inspecting Chain-of-Thought reasoning is among the most common means of understanding why an LLM produced its output. But well-known problems with C

Sources: officials say OpenAI risks its White House relationship by hiring Dean Ball, who has criticized Trump's AI strategy after leaving the administration (Thomas Barrabi/New York Post)

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
SafetyDGX agent

Thomas Barrabi / New York Post: Sources: officials say OpenAI risks its White House relationship by hiring Dean Ball, who has criticized Trump's AI strategy after leaving the administration — Top Open

The Horizon Gap: Planning, Memory, Execution, Training, and Evaluation for Long-Horizon LLM Agents

SafetyDGX agent

arXiv:2608.06663v1 Announce Type: new Abstract: Frontier language models solve reasoning problems in a single forward pass that would have been research contributions years ago, yet fail at multi-hour

Very cool idea to have agents design complex systems by searching over the model structure itself. New research from Sakana AI introduces CE…

Model ReleasesDGX agent

Very cool idea to have agents design complex systems by searching over the model structure itself. New research from Sakana AI introduces CEDAR, which uses LLM agents to write, simulate, and refine sy

Why Speculative Decoding went mature in 2026?

Local AiDGX agent

Spec-dec has been a thing for a while, in fact, it's wasn't an idea that was born for LLM inference. E.g. Uber's https://github.com/uber/submitqueue applied it to a merge queue. Apple & GDM had been r

9 Aug 2026

[2606.05682] Beyond Output Matching: Preserving Internal Geometry in NVFP4 LLM Distillation

Model ReleasesDGX agent

Demand for low-precision inference, including NVFP4-based approaches, has grown as large language models are increasingly deployed in latency and cost constrained production environments. Quantization

KLQ: Training-free measured rotation quantization. Beats all training-free rotation-based quantization methods on W4A4KV4-bits. Llama 3.2 1B KLQ-quantized beats SpinQuant and gets close to ReSpinQuant without GPTQ/LDLQ rounding.

Model ReleasesDGX agent

First of all, I'm not a lab, this was a solo summer research project that finally culminated into the github repo and the writeup. The repo includes a much deeper dive with methods, findings about qua

7 Aug 2026

Agentic Nesting: A New Methodology for Existing Enterprise Application Integration and Services

AgentsDGX agent

arXiv:2608.05159v1 Announce Type: new Abstract: Enterprise operations extensively rely on multiple heterogeneous business systems and information applications, which also result in severe data silos a

At Black Hat, OpenAI reconstructs the OpenAI-Hugging Face incident and examines its implications for AI security, cyber resilience, and alignment (Black Hat on YouTube)

SafetyDGX agent

Black Hat on YouTube: At Black Hat, OpenAI reconstructs the OpenAI-Hugging Face incident and examines its implications for AI security, cyber resilience, and alignment — The ‘Breaking’ News: The OpenA

Autonomous Research Agents: A Survey of AI Scientists and the Verification Gap

Model ReleasesDGX agent

arXiv:2608.05179v1 Announce Type: cross Abstract: Large language model (LLM) agents are increasingly used across the scientific research lifecycle: ideation, literature search, experiment design and e

Berkeley and Heiserman as an Unexhausted Architecture for Embodied Machine Intelligence

ResearchDGX agent

arXiv:2607.16465v2 Announce Type: replace Abstract: Edmund C. Berkeley is usually remembered as a mediator between symbolic logic and early computing, yet that standard description understates the sco

Every CIO should watch this. Persistent systems trying to break through and solve problems at all costs are more creative than you think. La…

IndustryDGX agent

Every CIO should watch this. Persistent systems trying to break through and solve problems at all costs are more creative than you think. Labs and enterprises will be focused more on network effects (

PaCoNet: Deep Data Extraction for Parallel Coordinates

ResearchDGX agent

arXiv:2608.06030v1 Announce Type: new Abstract: Extracting data from visualizations has long challenged computer vision, with current research focused on bar, line, and pie charts, among other low-dim

Text Generation: A Systematic Literature Review of Tasks, Evaluation, and Challenges

SafetyDGX agent

arXiv:2405.15604v4 Announce Type: replace Abstract: Text generation has become more accessible than ever, and the growing interest in these systems, especially those using large language models, has s

Vibe Compiler: A Research-Logic Synthesis Tool That Runs without Prompt Engineering -Toward Enhancing Metacognition for Sustaining Agency in the Age of Generative AI-

Model ReleasesDGX agent

arXiv:2608.05545v1 Announce Type: cross Abstract: Generative AI used as a capable servant has greatly accelerated intellectual work, but it also risks eroding human epistemic agency by encouraging unc

6 Aug 2026

A Long-Run Persistence Theory for AI Systems under the Redundancy-Adjusted Artificial Age Score (AAS)

ResearchDGX agent

arXiv:2608.04012v1 Announce Type: new Abstract: Artificial intelligence systems are increasingly expected to operate over repeated cycles of interaction, adaptation, and update rather than through iso

Fundamentals of quantum Boltzmann machine learning with visible and hidden units

ResearchDGX agent

arXiv:2512.19819v2 Announce Type: replace-cross Abstract: One of the primary applications of classical Boltzmann machines is generative modeling, wherein the goal is to tune the parameters of a model

Short-term load forecasting under EU-AI Act Requirements in Safety-Critical Environments: Results from a 41-day live challenge on the aggregated German transmission-grid load

Model ReleasesDGX agent

arXiv:2608.05018v1 Announce Type: new Abstract: Short-term load forecasting (STLF) play a vital role in the electric power industry. It serves infrastructure that European and German law designate as

5 Aug 2026

A New Theory of Value for Post-AGI Economics

TutorialsDGX agent

arXiv:2608.01432v2 Announce Type: replace Abstract: Artificial general intelligence (AGI) may weaken scarcities in labour, expertise, information, and productive capability that underpin established t

Hear to See: Discerning Stateful Listening for Audio-Visual Instance Segmentation

ResearchDGX agent

arXiv:2608.03264v1 Announce Type: cross Abstract: Audio-visual instance segmentation (AVIS) requires accurately identifying and tracking individual sounding objects with pixel-level masks. Existing me

Predictive Set Theory: A Generative Framework for Cognitive Architecture with Operationalized Core Mechanisms

ResearchDGX agent

arXiv:2608.02704v1 Announce Type: new Abstract: Predictive processing theories portray the brain as a hierarchical prediction engine that minimizes prediction error, yet they lack operational definiti

The Agent Operating System (AOS): A Reference Operating Architecture for Distributed Agentic Systems

SafetyDGX agent

arXiv:2608.03214v1 Announce Type: new Abstract: Large language models have transformed artificial intelligence from isolated prediction services into components of long-running, distributed systems th

4 Aug 2026

Building agents that patch other agents. This is an interesting approach for self-improving agents that leverages agent outputs. If you run …

AgentsDGX agent

Building agents that patch other agents. This is an interesting approach for self-improving agents that leverages agent outputs. If you run agents in production, you already have the training data for

From Pixels to PCells: A Neurosymbolic Approach to Photonic Component Creation

ResearchDGX agent

arXiv:2608.00084v1 Announce Type: new Abstract: We present PixCell, a neurosymbolic system in which multimodal agents convert a visually presented photonic component into a parametric program over a s

PackingGPT: 3D Packing Agent for Real Furniture in Last-Mile Delivery

Model ReleasesDGX agent

arXiv:2608.01427v1 Announce Type: new Abstract: 3D bin packing rectangular items into standardised containers to maximise space utilisation under geometric shipping automation. Loading a furniture pur

Picking the right agent harness is now a crucial skill for any AI engineer. Imagine using the same model, same task, and same prompt. Now mo…

Model ReleasesDGX agent

Picking the right agent harness is now a crucial skill for any AI engineer. Imagine using the same model, same task, and same prompt. Now move it between two agent harnesses and the cost per success c

TrimMoE A communication aware and adaptive depth framework for distributed edge inference

Model ReleasesDGX agent

arXiv:2608.00573v1 Announce Type: cross Abstract: Serving Mixture-of-Experts (MoE) large language models across distributed edge servers is bottlenecked by the cross-server expert transmission. The ex

3 Aug 2026

A user's guide to PINNs in geometric analysis: lessons from the asymptotic Plateau problem

TutorialsDGX agent

arXiv:2607.28733v1 Announce Type: cross Abstract: This proceedings contribution elaborates on the findings of arXiv:2605.26234v2: a joint work with Marco Usula, where we introduced a machine learning

Beyond Component Testing: Validating Agentic AI Systems

SafetyDGX agent

arXiv:2607.29405v1 Announce Type: new Abstract: Agentic AI systems act through multi-step trajectories that combine planning, tool use, memory, interaction, and adaptation. This behavior stretches val

Nice benchmark to measure agentic e-commerce capabilities. They ran an agent for one simulated year of e-commerce operations and it ends up …

Model ReleasesDGX agent

Nice benchmark to measure agentic e-commerce capabilities. They ran an agent for one simulated year of e-commerce operations and it ends up with 27.3% of the money a human makes. MerchantBench is a 36

The Theoretical Foundation of Socratic Tests: Dynamic, Multimodal, Conversational Examinations

SafetyDGX agent

arXiv:2607.29624v1 Announce Type: cross Abstract: Traditional static assessments rely on a subtractive, deficit-based grading model that often penalizes ambition and obscures diagnostic feedback. Conv

2 Aug 2026

[Release] WinterMix — Qwen3.5-122B-A10B in native MLX: an 82 GiB build that beats 94–95 GiB quants, plus a 68 GiB build for agent swarms

Model ReleasesDGX agent

TL;DR: I spent 9 days developing a new quantization method for MLX models and measured 18 variants against each other on a single M5 Max MacBook Pro (128 GB). The result is the best-measuring MLX quan

Vacuum 16T

Model ReleasesDGX agent

https://huggingface.co/tsfrm/vacuum-16t A 16.5-trillion-parameter model that contains nothing. This model is just a ████ you to the labs and companies who say that 'haha I have the biggest model out t

1 Aug 2026

3/ here's the part that makes it non-optional: the same agent that will do whatever it takes to solve a problem will also walk straight out …

Model ReleasesDGX agent

3/ here's the part that makes it non-optional: the same agent that will do whatever it takes to solve a problem will also walk straight out of a sandbox you thought was locked down. We watched exactly

// Persistent Workspaces for Long-Lived Claude Code Agent Teams // Four issues to be aware of: > Working state vanishes when a terminal clos…

Model ReleasesDGX agent

// Persistent Workspaces for Long-Lived Claude Code Agent Teams // Four issues to be aware of: > Working state vanishes when a terminal closes and the team cannot be resumed. > Compaction condenses th

The “AGI-is-near” community keeps committing the same logical fallacy over and over; I have seen it at least half a dozen times today alone.…

SafetyDGX agent

The “AGI-is-near” community keeps committing the same logical fallacy over and over; I have seen it at least half a dozen times today alone. Every time there’s an advance, I see the same error. Here’s

31 Jul 2026

AI-assisted pre-review of open-source software submissions: an experience report from BOSC 2026

SafetyDGX agent

arXiv:2607.27228v1 Announce Type: new Abstract: Most conferences rely on peer-review of submissions, but as generative AI makes it easier than ever to prepare submission materials, some conferences ar

AI Security Priorities: A Field-Wide Agenda

SafetyDGX agent

arXiv:2607.26069v1 Announce Type: cross Abstract: As AI systems are rapidly integrated into critical economic, governmental, and national security functions, the gap between AI adoption and AI securit

Behavioral Controllability of Agentic Models for Information Extraction: From Fixed Workflows to Reflective Agents

AgentsDGX agent

arXiv:2607.15715v2 Announce Type: replace Abstract: Large language model (LLM) agents are increasingly used for complex information-extraction tasks, yet it remains unclear whether agentic components

Direct Rotor Thrust Sensing and Feedback Control for Disturbance Rejection of Multirotors Using Load-cells

ResearchDGX agent

arXiv:2607.10099v2 Announce Type: replace Abstract: Gust disturbances, dynamic vertical inflow and ground effect are key adverse aerodynamic phenomena that induce variations in the forces acting on a

Neat work on long-horizon agents. Splitting a hard task across agents is typically how standard multi-agent work. The usual design lets them…

Model ReleasesDGX agent

Neat work on long-horizon agents. Splitting a hard task across agents is typically how standard multi-agent work. The usual design lets them exchange findings only at phase boundaries, through staged

New research from Microsoft. This one is on training computer-use agents at scale. Recent pipelines generate synthetic environments in bulk,…

Model ReleasesDGX agent

New research from Microsoft. This one is on training computer-use agents at scale. Recent pipelines generate synthetic environments in bulk, which moved the bottleneck from how many exist to what is i

Pramana: A Composable, Domain-Specific Backend for Empirical Networking Research

AgentsDGX agent

arXiv:2607.26352v1 Announce Type: cross Abstract: Networking research advances by turning hypotheses into empirical evidence, so accelerating it means reducing the lag between ideation (synthesizing a

SciSchema.org: A Multidisciplinary Collection of Schemas for Structured Scientific Process Descriptions

ResearchDGX agent

arXiv:2607.27955v1 Announce Type: cross Abstract: Scientific processes are often described in heterogeneous article discourse, with details needed for comparison, reproducibility, reuse, and automatio

30 Jul 2026

Actions Have Consequences: Detecting Outcome Performativity using Intervention Testing

ApplicationsDGX agent

arXiv:2607.26908v1 Announce Type: new Abstract: In many domains such as Palliative Care, Credit Assignment and Recommender Systems, predictions may causally influence the outcomes they predict. This p

As the benchmarks that test frontier AI on get more complex, we are losing one of the most important aspects of benchmarking: comparisons to…

ApplicationsDGX agent

As the benchmarks that test frontier AI on get more complex, we are losing one of the most important aspects of benchmarking: comparisons to humans Validated benchmarks need to have human (ideally mul

Automorphism-Induced Non-Canonicity in Top-k Explanations of Graph Neural Networks

Model ReleasesDGX agent

arXiv:2607.26344v1 Announce Type: new Abstract: A gradient-based GNN explainer given a molecule with two chemically equivalent nitro groups assigns them attribution scores that are equal to the last b

Multi-Objective Compliance-Integrated Coevolution For Simulated And Real-World Deployment Of Multi-Robot Marine Autonomy

SafetyDGX agent

arXiv:2607.26279v1 Announce Type: new Abstract: Collaborative robots are well-suited to maritime missions that benefit from coordination, such as the exploration of unknown reef structures, inspection

Online Handwriting Trajectory Reconstruction from Kinematic Sensors using Temporal Convolutional Network

Model ReleasesDGX agent

arXiv:2607.26733v1 Announce Type: new Abstract: Handwriting with digital pens is a common way to facilitate human-computer interaction through the use of Online Handwriting (OH) trajectory reconstruct

Scientific Knowledge Discovery in the Age of Large Language Models

ResearchDGX agent

arXiv:2607.26670v1 Announce Type: cross Abstract: The rapid growth of scholarly literature has made identifying relevant publications increasingly difficult, and conventional search systems still depe

SciFigQual-Bench: A Benchmark for Scientific Figure Quality Assessment with Full-Manuscript Context

Model ReleasesDGX agent

arXiv:2607.27084v1 Announce Type: new Abstract: Scientific images are the core elements of presenting experimental conclusions, elaborating system architecture, and supporting comparative arguments in

We have to be careful to not offload our understanding to agents. I think there is also a good opportunity to build agentic applications tha…

AgentsDGX agent

We have to be careful to not offload our understanding to agents. I think there is also a good opportunity to build agentic applications that encourage deeper understanding. For example, coding agents

When benchmark inferences do not compose: Projectibility in AI evaluation

Model ReleasesDGX agent

arXiv:2607.26159v1 Announce Type: cross Abstract: An AI benchmark result rarely reaches a consequential claim in one step. Evaluators generalize it to further cases, interpret it as evidence of capabi

29 Jul 2026

AI Deployment and Cyber Governance Failures in Public-Sector Organizations: A Typological Analysis

ResearchDGX agent

arXiv:2607.25368v1 Announce Type: new Abstract: The intersection of artificial intelligence adoption, cybersecurity governance, and public sector institutional constraints has not been examined as a u

Context Assembly as the Controlled Variable: A Control-Theoretic View of Harness Policies for Frozen LLM Agents

SafetyDGX agent

arXiv:2607.25408v1 Announce Type: new Abstract: A growing body of 2026 work applies control theory to LLM agents: Lyapunov-certified stability for tool-mediated controllers (Prinos et al., 'Stable Age

Decision-Level Hijacking: Injecting Cognitive Bias into Large Language Models via Bit-Flip Attacks

Model ReleasesDGX agent

arXiv:2607.25227v1 Announce Type: cross Abstract: Large Language Models (LLMs) have been widely applied in high-stakes decision-making scenarios such as corporate strategy, and users are increasingly

F(AI)2R: Who Did What, and Who Checked? Verifiable AI Provenance as an Executable Skill

AgentsDGX agent

arXiv:2607.25637v1 Announce Type: cross Abstract: F(AI)2R is FAIR research with AI in the loop, twice: an AI-assisted authoring pass and a machine-readable audit pass over every artefact. AI systems n

Finding Optimal Cost-Bounded Plan Reductions: Refined Model

ResearchDGX agent

arXiv:2607.25484v1 Announce Type: new Abstract: In some real applications a plan may later become unfeasible due to newly imposed budget constraints, yet, at the same time, using only the original act

Forgetting Is Not a Fix: Path Dependence in Sequential Engram Editing

ResearchDGX agent

arXiv:2607.24805v1 Announce Type: cross Abstract: AI Engram (Kwon et al., 2026) formalizes the four engram criteria of neuroscience as a constrained inverse problem in weight space and solves it close

← Previous
1…4142434445…203
Next →