AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
83,860 results
7 Aug 2026

Where Models Converge and Humans Diverge: A Coverage Framework for Distributional Pluralism in Open-Ended Generation

ResearchDGX agent

arXiv:2608.05576v1 Announce Type: new Abstract: When a large language model (LLM) writes Harry Potter fanfiction, it reliably produces fundamental elements of the Hogwarts universe, such as recognizab

Where Privacy Risk Lives in English-Source Multilingual RAG: A Stage-Decomposed Audit Across Five Query Languages

Model ReleasesDGX agent

arXiv:2608.05163v1 Announce Type: new Abstract: A common assumption holds that switching to a non-English language makes a multilingual RAG system easier to attack for personal information. We test th

Who Checks the Citations? Benchmarking Legal Hallucination Detection

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub

arXiv:2606.21155v2 Announce Type: replace Abstract: Attorneys, judges, and pro se filers increasingly use AI to draft legal documents, yet these tools frequently fabricate citations. Despite predictio

Who Gets Access? Global Region and Academic Status Bias in AI-Generated Academic Gatekeeping Scenarios

SafetyDGX agent

arXiv:2608.05178v1 Announce Type: cross Abstract: Equitable access to scientific knowledge often depends on informal gatekeeping decisions, particularly when resources such as paywalled articles, data

Why the Third Axis Is Freedom

ResearchDGX agent

arXiv:2608.05423v1 Announce Type: cross Abstract: In generative training, a model produces an output and is penalised for its difference from an example. With one output per comparison, a model that p

With agentic AI, workflows are increasingly CPU hungry. The share of cognition moving to the CPU keeps increasing.

AgentsDGX agent

With agentic AI, workflows are increasingly CPU hungry. The share of cognition moving to the CPU keeps increasing. Scoop: AWS engineers have been told to conserve CPU compute to make sure the cloud gi

Woodpecker Distillation: Weak Models Diagnose Reasoning Bugs in Strong Models

Local AiDGX agent

arXiv:2608.05168v1 Announce Type: new Abstract: Large language models often fail on reasoning tasks despite possessing the capability to solve them. We argue that many such failures arise from localiz

World-to-Wrist: Task-Conditioned Future Wrist Modeling for Fine-Grained Robot Manipulation

Local AiDGX agent

arXiv:2608.05369v1 Announce Type: cross Abstract: Vision-language-action (VLA) models often treat main-view and wrist-view observations as parallel visual inputs, overlooking their distinct roles in r

WorldClaw: Agentic 3D Open-World Generation at Scale

AgentsDGX agent

arXiv:2608.05248v1 Announce Type: new Abstract: Generating large-scale, freely explorable 3D worlds from open-ended text remains challenging because a system must jointly maintain global spatial coher

Worst-Case Distance-Aware Error Bounds for Neural Networks

SafetyDGX agent

arXiv:2510.22021v3 Announce Type: replace Abstract: Safety-critical applications of machine learning require uncertainty estimates that support reliable worst-case analysis. Neural networks (NNs) prov

XEWorld: Can Action-Conditioned World Models Generalize to Unseen Robot Embodiments?

SafetyDGX agent

arXiv:2608.05799v1 Announce Type: cross Abstract: Action-conditioned world models are promising learned simulators for robotic manipulation, yet evaluating them exclusively on training robots fails to

You can play it here: https://simonw.github.io/raccoon-heist-codex/ For comparison, here's Fable 5 + Claude Code's game, built from the exac…

Model ReleasesDGX agent

You can play it here: https://simonw.github.io/raccoon-heist-codex/ For comparison, here's Fable 5 + Claude Code's game, built from the exact same prompt https://x.com/simonw/status/208508951822360205

Zero-code, low-cost data ingestion: New BigQuery DTS capabilities

AgentsDGX agent

In a fast-paced digital economy, data is your most critical engine. Yet, many enterprises find themselves trapped in a costly paradox, spending over 100 hours a week building and fixing fragile, in-ho

Zero-Shot Multi-Disease Labeling of Chest, Abdomen, and Pelvis CT Reports Using Open-Weight Large Language Models: The Effect of Labeling Conventions

Model ReleasesDGX agent

arXiv:2506.03259v3 Announce Type: replace Abstract: Purpose: To compare five lightweight open-weight large language models (LLMs) with a rule-based algorithm (RBA) and fine-tuned RadBERT for zero-shot

6 Aug 2026

2 x 5070ti Qwen 27B full config / stats

Model ReleasesDGX agent

Following up on yesterday's post about running everyone's faves on 2 x 16gb cards while maximizing performance and KV. Previous post data used abandoned Cu130 VLLM image. Stats here are done on cu129-

5.6 Sol much better in chat now and unlimited text chat for free users!

Model ReleasesDGX agent

5.6 Sol much better in chat now and unlimited text chat for free users! We’re making better intelligence easier to access in ChatGPT for everyone: - GPT-5.6 Sol now powers both Instant and deep reason

A 6G Integrated Sensing and Communication Framework for Railway Intrusion Detection and Collision Prediction

ResearchDGX agent

arXiv:2608.04710v1 Announce Type: cross Abstract: Integrated Sensing and Communication (ISAC) combines sensing and communication to efficiently utilize wireless resources and is emerging as a key para

A Chain Is Only as Strong as Its Weakest Link: A Scoping Review of System Integration Audits in AI

SafetyDGX agent

arXiv:2608.04921v1 Announce Type: cross Abstract: As AI systems become increasingly integrated into diverse interfaces and applications, model-centric audits are insufficient to address risks arising

A Comparative Study of Feature Selection Methods for EHR Diagnosis Codes in Opioid Use Disorder Prediction

ResearchDGX agent

arXiv:2608.04180v1 Announce Type: new Abstract: Feature selection is a critical step in electronic health record (EHR)-based predictive modeling, where input variables are often high-dimensional, spar

A Counterexample to Fourier Alignment in Single-Neuron Modular Addition

Model ReleasesDGX agent

arXiv:2608.04451v1 Announce Type: cross Abstract: We give a negative solution to MAIS-O60. We first construct an example in which an initially active ReLU neuron becomes completely inactive in finite

A General Sufficient Condition for Rewriting Horn-ALCHI Atomic Queries into GQL

ResearchDGX agent

arXiv:2608.04945v1 Announce Type: cross Abstract: The emergence of the ISO standard GQL introduces a powerful query language extending first-order logic with controlled recursion, raising the question

A geometry-based deep equilibrium model for image restoration under multiplicative Gamma noise

ResearchDGX agent

arXiv:2608.04944v1 Announce Type: cross Abstract: We propose a deep learning framework for image restoration from images degraded by both multiplicative Gamma noise and blur. Unlike conventional deep

A GitOps-Driven Annotation Catalog for Fully Automatic Railway Operations

ApplicationsDGX agent

arXiv:2608.04724v1 Announce Type: cross Abstract: Automatic train operation (ATO) at grade of automation 3 and above (GoA3-GoA4) requires robust AI-based perception systems capable of reliably detecti

A guide to slash commands in the GitHub Copilot app

TutorialsDGX agent

Go beyond chat in the GitHub Copilot app with these slash commands. They'll help you plan, collaborate, automate, and customize your dev workflow. The post A guide to slash commands in the GitHub Copi

A Long-Run Persistence Theory for AI Systems under the Redundancy-Adjusted Artificial Age Score (AAS)

ResearchDGX agent

arXiv:2608.04012v1 Announce Type: new Abstract: Artificial intelligence systems are increasingly expected to operate over repeated cycles of interaction, adaptation, and update rather than through iso

A look at Sequoia's revamped strategy under new stewards Alfred Lin and Pat Grady, including bold AI bets; sources: Sequoia recently closed $10B in new funding (Bloomberg)

IndustryDGX agent

Bloomberg: A look at Sequoia's revamped strategy under new stewards Alfred Lin and Pat Grady, including bold AI bets; sources: Sequoia recently closed $10B in new funding — Alfred Lin and Pat Grady ar

A Mechanistic Analysis of Transformers for Dynamical Systems

ResearchDGX agent

arXiv:2512.21113v2 Announce Type: replace Abstract: Transformers are increasingly adopted for modeling and forecasting time-series, yet their internal mechanisms remain poorly understood from a dynami

A Model Merging Approach for Continual MLLM Unlearning

ResearchDGX agent

arXiv:2608.04548v1 Announce Type: cross Abstract: Multimodal large language model (MLLM) unlearning methods have been proposed to remove private, sensitive, or proprietary information from well-traine

A Modular Part-of-Speech Tagger for Scottish Gaelic using spaCy

ResearchDGX agent

arXiv:2608.04808v1 Announce Type: new Abstract: Part-of-speech tagging for low-resource languages remains challenging due to limited annotated data, especially for linguistically complex languages. Ga

A Multi-Cohort Validation of Censoring-Aware Conformal Lower Predictive Bounds for Pathology Survival Models

ResearchDGX agent

arXiv:2608.04025v1 Announce Type: cross Abstract: Whole-slide survival models commonly provide risk rankings without calibrated statements about individual event times. We evaluate fixed-cutoff drcosa

A Multi-Sensor Dataset for Monitoring the Operational Environment of Rail Vehicles

ResearchDGX agent

arXiv:2608.04704v1 Announce Type: new Abstract: Reliable environment monitoring is essential for the safe and efficient operation of automated railway systems, covering all Grades of Automation (GoA),

A quick snapshot of where Qwen3.8-Max stands today: Qwen3.8-Max now ranks #5 on the Artificial Analysis Intelligence Index, and #1 on the Ag…

AgentsDGX agent

Qwen 3.8‑Max, Alibaba’s latest large language model, was reported on August 6 2026 to rank **#5** on the Artificial Analysis Intelligence Index and **#1** on the Agentic Index. These rankings position

A-SR: Self-Evolving Agentic LLMs for Symbolic Regression via Hierarchical Coordination

SafetyDGX agent

arXiv:2608.04872v1 Announce Type: cross Abstract: Symbolic regression aims to discover closed-form equations from data, but existing LLM-guided methods often rely on a unified proposal loop that compr

A Survey of Agent Memory in the Second Half: Towards Self-Evolving and Long-Horizon Agents

Model ReleasesDGX agent

arXiv:2602.06052v4 Announce Type: replace-cross Abstract: Research in artificial intelligence is shifting from model innovations and benchmark scores towards problem definition and rigorous real-world

A Trust-region Framework for Moment Estimation

ResearchDGX agent

arXiv:2608.04026v1 Announce Type: cross Abstract: In this paper, we develop a trust-region framework for understanding the behavior of adaptive moment estimation mechanisms, such as extsc{Adam}, in st

A Unified Model for Cross-Domain Clone Detection via Model Merging

Model ReleasesDGX agent

arXiv:2608.04215v1 Announce Type: cross Abstract: The growing diversity of code clone types, from syntactic copies to cross-language semantic clones to AI-generated duplicates, has created a fragmenta

A Vision-based Control Framework for Real-time Autonomous UUV Operations

AgentsDGX agent

arXiv:2608.04723v1 Announce Type: new Abstract: This paper presents a fully integrated vision-based framework for real-time and robust localization, autonomous navigation, and mapping for unmanned und

A/B Agent: A Self-Evolving Agent for Strategy Iteration in Industrial A/B Testing

SafetyDGX agent

arXiv:2608.04625v1 Announce Type: new Abstract: Industrial recommendation strategy iteration heavily relies on large-scale A/B experimentation. Traditional tuning requires experts to repeatedly design

Above-ground Biomass Estimation with Geospatial Foundation Models

Model ReleasesDGX agent

arXiv:2608.04792v1 Announce Type: new Abstract: Accurate estimation of Above-Ground Biomass (AGB) from satellite imagery is essential for the large-scale monitoring of carbon stocks, yet it remains a

ABSeeker: Training Long-Horizon Search Agents via Answer-Backtracked Credit Assignment

ResearchDGX agent

arXiv:2608.05102v1 Announce Type: new Abstract: Long-horizon search agents must make multiple sequential actions (steps) to search, retrieve, verify, and integrate evidence to reach a final answer. Ho

ACA-GS: Adaptive-Capacity Anchored Gaussian Splatting for Compact Dynamic Radiance Fields

Model ReleasesDGX agent

arXiv:2608.04581v1 Announce Type: new Abstract: Recent advances in 4D Gaussian Splatting (4DGS) enable high-fidelity, real-time spatiotemporal rendering, but expose a fundamental trade-off between mot

Active Learning Guided Design Space Refinement for Scalable Multi-Objective Bayesian Optimization in Materials Discovery

AgentsDGX agent

arXiv:2608.04651v1 Announce Type: new Abstract: Advanced materials discovery increasingly relies on machine learning and Bayesian optimization to explore large discrete design spaces under limited eva

Active-SWE: Benchmarking Coding Agents for Proactive Bug Fixing without Issue Reports

Model ReleasesDGX agent

arXiv:2608.04682v1 Announce Type: cross Abstract: Coding agents powered by large language models (LLMs) are increasingly adopted in software engineering (SWE) scenarios, capable of fixing a specific b

Actual is a local inference stack optimized to let you utilize your personal compute. With its low CPU impact while inferencing and multi-pl…

Local AiDGX agent

Actual is a local inference stack optimized to let you utilize your personal compute. With its low CPU impact while inferencing and multi-platform portability, it’s a great pair for your local Hermes

Adaptive Finite-Budget Training for CVaR Risk-Aware Q-Learning

Model ReleasesDGX agent

arXiv:2608.04305v1 Announce Type: new Abstract: Risk-aware Q-learning (RaQL) provides a model-free, two-timescale estimator for dynamic risk objectives, but its finite-budget behavior remains fragile:

Advancing brain tumor research with privacy-first AI

Model ReleasesDGX agent

The intersection of medicine and AI has led to remarkable innovations. However, developers now face the thorny challenge of building robust medical AI tools that have been tested and evaluated on dive

Advancing Utility Pole and Sign Detection Through Deep Learning

Model ReleasesDGX agent

arXiv:2608.04061v1 Announce Type: new Abstract: Utility poles are an essential part of the infrastructure used to support power distribution systems and other critical public services. Their regular i

Adversarial Attacks for Good: A Survey of Proactive Protection across the Visual Content Lifecycle

Model ReleasesDGX agent

arXiv:2608.04314v1 Announce Type: cross Abstract: Once visual content enters an AI pipeline, its owner often retains little technical control over how it is used. Legal and regulatory remedies can add

Adversarially Robust Abductive Fusion of Pre-trained Transformer-based Perception Models

Model ReleasesDGX agent

arXiv:2608.04190v1 Announce Type: new Abstract: Deploying pre-trained perception models in novel environments degrades their accuracy under distributional shift, and assembling them alone does not rec

AFD-Ledger: Deployment Provisioning for Attention--FFN Disaggregation

ResearchDGX agent

arXiv:2608.04502v1 Announce Type: cross Abstract: Attention--Feed-Forward Network (FFN) Disaggregation (AFD) is emerging as a promising architecture for serving Mixture-of-Experts (MoE) language model

Agent Skills for Automated Reasoning policies in Amazon Bedrock

SafetyDGX agent

Learn how to run the full Amazon Bedrock Automated Reasoning policy lifecycle from your coding agent. A suite of open source Agent Skills builds, reviews, tests, debugs, deploys, and validates a custo

AgentAntibody: An Adaptive Immune System for Defending LLM Agents against Prompt Injection

AgentsDGX agent

arXiv:2608.04053v1 Announce Type: cross Abstract: Prompt injection remains a critical threat to LLM agents, yet existing defenses treat each task as a self-contained problem, independent of previous e

AgentForge: An Immersive Role-Playing Platform for Learning Agentic Software Engineering

AgentsDGX agent

arXiv:2608.04148v1 Announce Type: cross Abstract: Agentic AI is increasingly used to coordinate planning, implementation, review, and testing in software development, yet it often offers limited trans

Agentic AI security tests enterprise defenses as scale outpaces strategy

AgentsDGX agent

Cybersecurity leaders are confronting an inflection point as agentic AI security becomes the defining challenge of this year’s threat landscape, with attackers and defenders racing to harness autonomo

Agentic Future Ready With BigQuery: Continually Improving Price-Performance, Zero Effort

Model ReleasesDGX agent

In the modern data landscape, query performance tuning and managing system price-performance is challenging, especially as the number of agentic workloads increase. Even for experienced developers and

Agentic Reinforcement Learning with Observation-Calibrated Self-Distillation

SafetyDGX agent

arXiv:2608.04788v1 Announce Type: cross Abstract: Large language model agents are commonly trained through reinforcement learning with sparse trajectory-level rewards, which offer limited guidance on

Agents want to collaborate. So we’re putting that instinct to work on improving open-weight LLMs at formal math. Our latest agent collab tac…

AgentsDGX agent

Agents want to collaborate. So we’re putting that instinct to work on improving open-weight LLMs at formal math. Our latest agent collab tackles the @SAIRfoundation challenge of building a cheat sheet

Agreement Before Diversity: Verification-First Complementarity for Heterogeneous Language-Model Coordination

ResearchDGX agent

arXiv:2608.04618v1 Announce Type: new Abstract: Heterogeneous language-model ensembles expand the space of candidate responses, yet they lack a principled criterion for when a newly generated answer s

AI agent observability: Why production systems need a reasoning layer

AgentsDGX agent

Traditional APM can collect every span and still leave developers guessing about intent, causality, and drift. As agents multiply, the observability stack must learn to interpret the systems it watche

AI-based single-shot structured-light depth reconstruction for real-time laparoscopic surgical guidance

HardwareDGX agent

arXiv:2608.05109v1 Announce Type: cross Abstract: Significance. Accurate intraoperative depth perception is important for autonomous and semi-autonomous robotic laparoscopic surgery. Conventional frin

← Previous
1…6970717273…1398
Next →