AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
10,030 results
5 Aug 2026

AI World Cup 2026: Benchmarking Large Language Models for End-to-End Football Tournament Prediction

Model ReleasesDGX agent

arXiv:2608.03416v1 Announce Type: new Abstract: Large language models (LLMs) are now regularly asked to forecast real-world events, but comparisons are often difficult because models receive different

b10289

Model ReleasesDGX agent

server: harden the file_glob_search directory walk (#26626) server: don't walk Windows junctions in file_glob_search std::filesystem reports a junction as a plain directory, so the symlink guard misse

Before Reasoning Can Fail: Pre-Evidence Procedural Failures in Agentic RAG

AgentsDGX agent

arXiv:2608.02011v2 Announce Type: replace Abstract: Agentic retrieval-augmented generation (RAG) systems can fail before evidence-conditioned reasoning is tested: an agent may retrieve candidate snipp

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Beyond the Single Camera: Agentic Multi-View Reasoning in Sports Video Understanding

Model ReleasesDGX agent

arXiv:2607.11844v2 Announce Type: replace Abstract: Recent Multimodal Large Language Models (MLLMs) achieve strong performance on single-view video understanding benchmarks. However, sports videos inv

Can LLM design high-quality experiments? A Comprehensive and Systematic Benchmark on Autonomous Experimental Design

Model ReleasesDGX agent

arXiv:2608.03501v1 Announce Type: new Abstract: AI for Research (AI4Research) leverages AI to automate and improve scientific workflows. While experimental design is a critical stage of the research p

Causal Inference with Unstructured Outcomes

TutorialsDGX agent

arXiv:2608.03085v1 Announce Type: cross Abstract: Causal inference has traditionally centered on scalar outcomes: whether a patient recovers, how much a worker earns, or how many visits a website rece

Cloudflare launches Cloudflare OS: an open-source AI agentic workspace for the enterprise

Model ReleasesDGX agent

Cloudflare Inc. today announced the launch of Cloudflare OS, an open-source artificial intelligence agentic workspace available through the browser, filled with custom shared micro-applications for en

CUADebug: Diagnosing and Repairing Computer-Use Agent Failures

Model ReleasesDGX agent

arXiv:2608.02643v1 Announce Type: cross Abstract: Computer-use agents (CUAs) operate real desktop and web interfaces through screenshots, mouse and keyboard actions, and stateful UI feedback, yet thei

Cura 1T: Specialized Model for Agentic Healthcare

Model ReleasesDGX agent

arXiv:2607.15314v2 Announce Type: replace Abstract: Healthcare AI agents handle patient consultation, clinical reasoning over text and images, interactive diagnosis, and electronic health record (EHR)

Designing Social Robots for Inclusive Child Wellbeing Assessment: Insights from Communities Supporting Developmental Language Disorder and Forced Migration

ResearchDGX agent

arXiv:2608.03820v1 Announce Type: new Abstract: Assessing children's wellbeing and mental health can be particularly challenging for children experiencing communication barriers, such as children with

Developers in Africa are increasingly choosing Chinese open-source AI models over US models, saying they are downloadable, easier to customize, and much cheaper (New York Times)

IndustryDGX agent

New York Times: Developers in Africa are increasingly choosing Chinese open-source AI models over US models, saying they are downloadable, easier to customize, and much cheaper — Developers built Sunf

Dr. AGENTONOMICS: A Didactic Experiment of AGENTONOMICS

AgentsDGX agent

arXiv:2608.03524v1 Announce Type: new Abstract: AGENTONOMICS is a framework that treats AI agents as economic entities that can be designed, managed, and governed through an integrated management arch

Enactive Artificial Intelligence: A Decision-Centric Architecture for Complex Systems

ApplicationsDGX agent

arXiv:2608.03413v1 Announce Type: new Abstract: As artificial intelligence (AI) continues to evolve and mature, recent AI practices have moved beyond large language models (LLMs) and text or image gen

From Social Coding to Agentic Coding: Productivity and Relational Reconfiguration in Open-Source Communities

Model ReleasesDGX agent

arXiv:2608.03585v1 Announce Type: new Abstract: Open-source software communities are a form of digital public infrastructure that not only produces code, but also generates public knowledge and interp

From Wearable Data to Personalized and Actionable Health Insights

ResearchDGX agent

arXiv:2608.03251v1 Announce Type: cross Abstract: Commercial wearable devices continuously capture rich physiological data (e.g., heart rate, respiration), opening new possibilities for monitoring hea

Information-Geometric Forward Policy Training in GFlowNets

Local AiDGX agent

arXiv:2608.03967v1 Announce Type: cross Abstract: Generative Flow Networks (GFlowNets) have emerged as a flexible framework for amortised inference over discrete and mixed discrete-continuous objects,

Interpreting Black-Box Large Language Models with Sentence-Level Energy Landscapes

Local AiDGX agent

arXiv:2608.02879v1 Announce Type: new Abstract: The widespread adoption of proprietary Large Language Models (LLMs) accessed strictly through closed APIs has created a critical challenge for responsib

Learning and Clustering on Temporal Graphs: Principles, Primitives, and Pooling

HardwareDGX agent

arXiv:2608.03696v1 Announce Type: new Abstract: This work focuses on the problem of learning on temporal graphs, with particular emphasis on the task of clustering: obtaining coarse-grained representa

Learning Clinical-Trial Strategy: Offline Policy Training for Decision Agents

SafetyDGX agent

arXiv:2608.03606v1 Announce Type: new Abstract: Clinical development is sequential decision-making under uncertainty, where a sponsor must plan a portfolio of experiments from heterogeneous evidence.

Measurement Without Validity: The Compounding Reliability Problem in Agentic AI Evaluation

Model ReleasesDGX agent

arXiv:2608.00794v2 Announce Type: replace Abstract: Agentic AI evaluation pipelines produce benchmark scores that justify deployment decisions, safety certifications, and regulatory compliance claims.

MultiGlobeQA: A Multilingual and Globally Diverse Benchmark for Geospatial Reasoning

Model ReleasesDGX agent

arXiv:2608.03882v1 Announce Type: cross Abstract: Geospatial reasoning, i.e., computing distances, containment, and other spatial relations over real-world entities, is central to navigation and logis

MVP-Tac: A Miniaturized Dual-Modal Vision and Photoelastic Tactile Sensor for Robot-Assisted Minimally Invasive Surgery

SafetyDGX agent

arXiv:2607.18660v2 Announce Type: replace Abstract: Robot-assisted minimally invasive surgery (RMIS) offers major benefits over open and conventional laparoscopic procedures, yet it still lacks tactil

Non-Destructive Quantification of Urea Adulteration in Bovine Milk Using Transmittance Multispectral Imaging

ResearchDGX agent

arXiv:2608.03113v1 Announce Type: new Abstract: Adulteration of bovine milk using urea remains a major food quality and health concern, motivating the development of rapid and quantitative screening t

None of this was us. Day 1 (and the first 48 hours) belonged to the open-source community: Generate with it @ComfyUI — native support + offi…

Local AiDGX agent

None of this was us. Day 1 (and the first 48 hours) belonged to the open-source community: Generate with it @ComfyUI — native support + official quantized builds, Day 0 Diffusers — the reference Pytho

One-shotting a Raccoon Heist game using Claude Fable 5

Model ReleasesDGX agent

Back in 2024 I tweeted screenshots of a game concept generated by GPT-3 and some concept 'art' created using DALL-E. Today, on the fourth anniversary of that tweet, I decided to see if Claude Fable 5

PAST-Bench: Benchmarking the Foundations of Recursive Self-Improvement in Personal Agents

Model ReleasesDGX agent

arXiv:2608.04003v1 Announce Type: new Abstract: Recursive self-improvement requires agents to turn accumulated experience into better future behavior. Personal AI agents offer a concrete setting for s

Prime Agent - a new coding harness surpassing Codex/CC/PI

Model ReleasesDGX agent

Prime Agent is an open-source coding and research agent for general and long-running work. A self-improving RLM harness for coding and long-running autonomous tasks. Designed to be both token-efficien

Principles of Robot Autonomy

AgentsDGX agent

arXiv:2608.03496v1 Announce Type: cross Abstract: Autonomous robots are moving rapidly from research labs into everyday life - on roads, in the air, in warehouses, and in space. Robot autonomy is no l

Run production AI agents in n8n with Amazon Bedrock AgentCore harness

AgentsDGX agent

Amazon Bedrock AgentCore harness is now generally available. Learn how to add it as an agent step in n8n workflows using a new open-source community node, and build agents with persistent memory, real

SAMSEM -- A Generic and Scalable Approach for IC Metal Line Segmentation

ApplicationsDGX agent

arXiv:2603.16548v2 Announce Type: replace-cross Abstract: In light of globalized hardware supply chains, the assurance of hardware components has gained significant interest, particularly in cryptogra

Security-First Evaluation of Text-to-Terraform: Benchmarking LLMs and SLMs for Secure IaC Generation

Model ReleasesDGX agent

arXiv:2608.02672v1 Announce Type: cross Abstract: Cloud misconfiguration remains a leading cause of security incidents, yet whether LLMs and SLMs can generate security-compliant Infrastructure-as-Code

Should We Type or Talk to LLM Agents? A Comprehensive Study of Voice and Keyboard Input Perturbations

ResearchDGX agent

arXiv:2608.03970v1 Announce Type: new Abstract: Human input reaches language models by typing or speaking, and each channel leaves a distinct signature: orthographic noise for keyboards; for voice, di

SimulRAG: Simulator-based RAG for Grounding LLMs in Long-form Scientific QA

Model ReleasesDGX agent

arXiv:2509.25459v2 Announce Type: replace Abstract: Large Language Models (LLMs) show promise in generating long-form scientific explanations that synthesize evidence and connect multiple factors. How

SMOPD: Multi-Reward Reinforcement Learning via Specialize-and-Merge Online Policy Distillation

SafetyDGX agent

arXiv:2608.03092v1 Announce Type: cross Abstract: We aim to improve model performance in multi-reward reinforcement learning training process. Existing Group reward-Decoupled Normalization Policy Opti

Studying, Identifying, and Fixing Hidden Technical Debt in AI-Intensive Cyber-Physical Systems

AgentsDGX agent

arXiv:2608.02638v1 Announce Type: cross Abstract: Artificial Intelligence (AI) components are increasingly pervasive in several software systems, including Cyber-Physical Systems (CPSs). AI-CPS are us

TabletCraft: Bridging a 4,000-Year Cultural Gap with Bidirectional Akkadian NMT and Cuneiform Rendering

ResearchDGX agent

arXiv:2608.02609v1 Announce Type: new Abstract: Half a million cuneiform clay tablets survive in museums worldwide, yet modern users can neither read nor write in the world's oldest writing system, le

Utilize a nvidia gpu and amd gpu together for 2 different ai models?

Model ReleasesDGX agent

We run a local model instance in our company that the dev we hired built for us. We're a trade business and we want to further use our on hand hardware for it. The specs given we have is a 5090 gpu wi

VIVID: A Culturally Grounded Benchmark Exposing the Figurative Language Gap in Vietnamese NLP

Model ReleasesDGX agent

arXiv:2608.03095v1 Announce Type: new Abstract: We present VIVID (Vietnamese Idioms for Validation and Interpretation Depth), the first systematic benchmark for evaluating culturally grounded figurati

4 Aug 2026

Augmented Inverse Hybrid Weighting: Robust Inference under Deterministic and Random Distribution Shifts

Model ReleasesDGX agent

arXiv:2608.00701v1 Announce Type: cross Abstract: Reweighting source samples to match a target covariate distribution is a standard response to distribution shift when generalizing evidence from one p

Development and Validation of a Dynamic Kidney Failure Prediction Model based on Deep Learning: A Real-World Study with External Validation

ApplicationsDGX agent

arXiv:2501.16388v3 Announce Type: replace Abstract: Background: Chronic kidney disease (CKD), a progressive disease with high morbidity and mortality, has become a significant global public health pro

Don't Offer What Can't Be Done: Deterministic Executability Gating for LLM Skill Selection at Scale

ApplicationsDGX agent

arXiv:2608.01050v1 Announce Type: cross Abstract: Production LLM agents that select from large skill libraries face a limitation that semantic relevance alone cannot resolve: a skill may match a user'

Emergence Invariance: From Symbolized Thought to Interface Refinement

Model ReleasesDGX agent

arXiv:2608.01548v1 Announce Type: cross Abstract: Language can be viewed as a formalized subset of thought: a consequence-governed symbolic structure projected from wider situated cognition. Large lan

Empowering Credit Risk Detection in Weixin Pay with Billion-Scale Deep Graph Learning

Local AiDGX agent

arXiv:2608.02168v1 Announce Type: new Abstract: Credit risk detection, particularly mitigating individual fraud, is crucial for maintaining the stability of digital financial ecosystems. Accurately id

Foundations of Reinforcement Learning and Control:Connections and New Perspectives

TutorialsDGX agent

arXiv:2608.02433v1 Announce Type: new Abstract: Reinforcement learning and control theory are two adjacent scientific fields that focus on optimizing the controller of unknown dynamical systems using

Given today, it is surprising how daring Microsoft & Google were initially with AI. Microsoft released GPT-4 before OpenAI, didn't back down…

Model ReleasesDGX agent

Given today, it is surprising how daring Microsoft & Google were initially with AI. Microsoft released GPT-4 before OpenAI, didn't back down after Sydney & got Copilot to market quickly (the 1st profe

GPT-OSS has turned one year old today!

Model ReleasesDGX agent

It is one of the best local models ever released, in both 20B and 120B versions. I always come back to it, especially the 120B version. Its only competition is, in my opinion, Qwen 3.5 122B, but that

How Deutsche Bank unlocked agility with an API-ready ecosystem

SafetyDGX agent

When people think about digital transformation in banking, they often focus on the visible results: mobile apps and new digital services. But there's an invisible infrastructure making all these servi

How Target is enhancing retail discovery and cutting database maintenance by 50% with Spanner Graph

SafetyDGX agent

In today’s retail environment, shoppers expect highly personalized product discovery experiences and conversational assistance that feels genuine, natural, and genuinely helpful. Today, successful pro

Information-Theoretic Foundations for Machine Learning

TutorialsDGX agent

arXiv:2407.12288v5 Announce Type: replace-cross Abstract: The progress of machine learning over the past decade is undeniable. In retrospect, it is both remarkable and unsettling that this progress wa

Leak It: A Probabilistic Approach to Training-Data Extraction from Black-Box Language Models

ResearchDGX agent

arXiv:2608.00144v1 Announce Type: cross Abstract: Membership inference (MIA) on language models is usually summarised by an aggregate ROC-AUC, but such evaluations are confounded: model-free blind bas

Learning Compositional Meta-Routing for Agentic Workflows: An Executable Benchmark

Model ReleasesDGX agent

arXiv:2608.00106v1 Announce Type: new Abstract: Agentic systems must decide not only what answer to produce, but which reasoning and execution operations should precede it. A controller may answer dir

LongChart VQA: A Comprehensive Benchmark for MLLMs with Complex Multi-Chart Reasoning

Model ReleasesDGX agent

arXiv:2608.01328v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) are rapidly evolving with expanded context windows and stronger reasoning capabilities, enabling multi-chart un

Multiple result sets: How Database Migration Service automates SQL server to PostgreSQL translation

Model ReleasesDGX agent

In the Medium blog post, 'From MARS to SETOF REFCURSOR: Migrating Multi-Result Stored Procedures to PostgreSQL,' we explored the fundamental architectural differences between SQL Server and PostgreSQL

onepot-Bench 0: towards lab-aware in silico chemistry benchmarks

Model ReleasesDGX agent

arXiv:2608.02595v1 Announce Type: new Abstract: Language models are playing an increasingly important role in laboratory science, performing tasks such as experiment planning, execution, and post-hoc

Open-DiffLoco: Open-Source Differentiable Learning for Deployable Blind Quadruped Locomotion

Local AiDGX agent

arXiv:2608.02069v1 Announce Type: cross Abstract: Developing deployable locomotion policies through conventional reinforcement learning often requires complex reward engineering and expensive training

Qwen-CUA: Native Computer Use for (almost) Everything

Model ReleasesDGX agent

arXiv:2608.02352v1 Announce Type: cross Abstract: Native computer use offers a general interface for agents to operate almost any software available to people, but requires long-horizon state tracking

Recursive Gaussian Processes and the Bayesian Brain

ResearchDGX agent

arXiv:2608.00503v1 Announce Type: cross Abstract: Predictive coding offers a powerful framework for cortical computation, yet scalable implementations that respect both Bayesian exactness and neurobio

Robust Bayesian Optimization via Tempered Posteriors

ResearchDGX agent

arXiv:2601.07094v2 Announce Type: replace-cross Abstract: Bayesian optimization (BO) iteratively fits a Gaussian process (GP) surrogate to accumulated evaluations and selects new queries via an acquis

Sampling-Based Visibility Task Planning

AgentsDGX agent

arXiv:2608.01027v1 Announce Type: new Abstract: Robot Task and Motion Planning (TAMP) algorithms enable autonomous operation by incorporating the specific functions and constraints of end-effector too

Scikit-fingerprints: Python library for scikit-learn compatible molecular fingerprints and chemoinformatics

TutorialsDGX agent

arXiv:2608.02027v1 Announce Type: new Abstract: We present scikit-fingerprints, a comprehensive, fully scikit-learn compatible library for molecular machine learning in Python, based on RDKit. Molecul

← Previous
1…106107108109110…168
Next →