AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “automated”

GridTimelineEvolution
4,983 results
Safety

RedDebate: Safer Responses Through Multi-Agent Red Teaming Debates

DGX agent

arXiv:2506.11083v3 Announce Type: replace Abstract: We introduce RedDebate, a novel multi-agent debate framework that provides the foundation for Large Language Models (LLMs) to identify and mitigate

safetyarxiv-cs-cl
2 Jun 2026
Model Releases

RenoBench: A Citation Parsing Benchmark

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2603.25640v2 Announce Type: replace-cross Abstract: Accurate parsing of citations is necessary for machine-readable scholarly infrastructure. But, despite sustained interest in this problem, exi

model-releasesarxiv-cs-cl
2 Jun 2026
Research

TalkTag: Fine-Grained Morphosyntactic Error Annotation for Transcribed Speech

DGX agent

arXiv:2606.01820v1 Announce Type: new Abstract: Fine-grained morphosyntactic error annotation is important in clinical and developmental language research, yet it is labour-intensive, expert-dependent

researcharxiv-cs-cl
2 Jun 2026
Agents

AutoSci: A Memory-Centric Agentic System for the Full Scientific Research Lifecycle

DGX agent

arXiv:2605.31468v1 Announce Type: new Abstract: Scientific research has traditionally been human-intensive, requiring researchers to coordinate literature, ideas, experiments, manuscripts, and review

agentsarxiv-cs-ai
1 Jun 2026
Agents

NEMO: Execution-Aware Optimization Modeling via Autonomous Coding Agents

DGX agent

arXiv:2601.21372v2 Announce Type: replace Abstract: We present NEMO, a system that translates Natural-language descriptions of decision problems into formal Executable Mathematical Optimization implem

agentsarxiv-cs-ai
1 Jun 2026
Safety

Organizational Adaptation to Generative AI in Cybersecurity

DGX agent

arXiv:2506.12060v2 Announce Type: replace-cross Abstract: Cybersecurity organizations are adapting to GenAI integration through modified frameworks and hybrid operational processes, with success influ

safetyarxiv-cs-ai
1 Jun 2026
Model Releases

SPM-Bench: Benchmarking Large Language Models for Scanning Probe Microscopy

DGX agent

arXiv:2602.22971v2 Announce Type: replace Abstract: As LLMs achieved breakthroughs in general reasoning, their proficiency in specialized scientific domains reveals pronounced gaps in existing benchma

model-releasesarxiv-cs-ai
1 Jun 2026
Safety

BEAMS: Benchmarking and Evaluating AI for Modeling and Simulation

DGX agent

arXiv:2605.28994v1 Announce Type: new Abstract: AI tools to support real world decision making must be able to build simulation models that inform their recommendations and render them interpretable.

safetyarxiv-cs-ai
29 May 2026
Agents

Croissant Tasks: A Metadata Format for Reproducible Machine Learning Evaluations

DGX agent

arXiv:2605.29786v1 Announce Type: new Abstract: Reproducibility is fundamental to the scientific method, yet remains a critical challenge in machine learning. Contributing factors include underspecifi

agentsarxiv-cs-ai
29 May 2026
Applications

Decentralized LLM-Driven Coordination of Acoustic Robots for Contactless Object Manipulation

DGX agent

arXiv:2605.29378v1 Announce Type: new Abstract: Natural language interfaces can simplify interaction with multi-robot systems, especially when non-expert users need to issue high-level commands. Acous

applicationsarxiv-cs-ro
29 May 2026
Tutorials

MEMENTO: Leveraging Web as a Learning Signal for Low-Data Domains

DGX agent

arXiv:2605.29795v1 Announce Type: new Abstract: Real-world tasks often lack large labeled datasets, motivating extensive work on learning in low-data regimes. However, existing approaches such as few-

tutorialsarxiv-cs-ai
29 May 2026
Applications

Motion-guided sparse correction enables expert-quality point tracking across diverse microscopy regimes

DGX agent

arXiv:2605.29220v1 Announce Type: new Abstract: Tracking the dynamics of non-canonical biological systems in microscopy videos remains a persistent challenge. Both classical and learning-based tracker

applicationsarxiv-cs-cv
29 May 2026
Model Releases

Announcing the newest cohort of the Google for Startups Accelerator: Middle East, North Africa & Turkey

DGX agent

Google’s mission is to organize the world’s information and make it universally accessible. In high-growth, technically ambitious markets like the Middle East, North Africa, and Türkiye (MENA-T), we f

model-releasesgoogle-cloud-ai
28 May 2026
Model Releases

CFDTwin: An open-source GUI and Python toolkit for POD-NN surrogate modeling of ANSYS Fluent simulations

DGX agent

arXiv:2605.27725v1 Announce Type: cross Abstract: High-fidelity computational fluid dynamics (CFD) is widely used for thermal-fluid design, but repeated CFD solves remain expensive for design optimiza

model-releasesarxiv-cs-lg
28 May 2026
Applications

REC-CBM: Rubric-Aware Error-Correction Concept Bottleneck Models for Trustworthy Open-Ended Grading

DGX agent

arXiv:2605.27402v1 Announce Type: cross Abstract: Open-ended grading is central to equitable and personalized education, yet manual grading remains time-consuming and costly, underscoring the need for

applicationsarxiv-cs-ai
28 May 2026
Research

The Shape of Reasoning: Topological Analysis of Reasoning Traces in Large Language Models

DGX agent

arXiv:2510.20665v3 Announce Type: replace Abstract: Evaluating the quality of reasoning traces from large language models remains understudied, labor-intensive, and unreliable: current practice relies

researcharxiv-cs-ai
28 May 2026
Safety

Beyond the Data Mesh Illusion: Designing Modern AI-augmented Lakehouses to Bridge the Gap Between Theory and Practice

DGX agent

arXiv:2605.27131v1 Announce Type: cross Abstract: Enterprise data platforms face an enduring tension between domain self-service and holistic governance. The data mesh paradigm proposed decentralized

safetyarxiv-cs-ai
27 May 2026
Model Releases

Strategies for Guiding LLMs to Use Software Design Patterns: A Case of Singleton

DGX agent

arXiv:2605.26898v1 Announce Type: cross Abstract: Large Language Models (LLMs) can generate functional source code from natural-language prompts, but often fail to consistently follow higher-level arc

model-releasesarxiv-cs-ai
27 May 2026
Research

Artificial Effort

DGX agent

arXiv:2605.23920v1 Announce Type: cross Abstract: Real-effort tasks, in which participants perform cognitively costly activities whose outcomes depend on actual performance, are widely used in experim

researcharxiv-cs-ai
26 May 2026
Model Releases

Beyond Summaries: Structure-Aware Labeling of Code Changes with Large Language Models

DGX agent

arXiv:2605.26100v1 Announce Type: cross Abstract: Code review is a critical practice in software engineering, yet the growing scale and frequency of code patches in modern projects, together with the

model-releasesarxiv-cs-ai
26 May 2026
Applications

DeIDClinic: A Risk-Aware Pseudonymization Framework for Clinical Text De-identification and Re-identification Risk Assessment

DGX agent

arXiv:2410.01648v2 Announce Type: replace Abstract: The increasing availability of sensitive textual data has created an urgent need for robust de-identification methods that enable compliant data sha

applicationsarxiv-cs-cl
26 May 2026
Model Releases

Double Triangle Annotation: A Scalable Human-in-the-Loop Framework for High-Precision Historical Document Annotation

DGX agent

arXiv:2605.25781v1 Announce Type: new Abstract: Evaluating structured-information extraction from historical documents at scale requires high-precision ground-truth annotations, yet traditional manual

model-releasesarxiv-cs-cl
26 May 2026
Applications

Knowledge Graph Re-engineering Along the Ontological Continuum (extended version)

DGX agent

arXiv:2605.22093v2 Announce Type: replace Abstract: Knowledge graphs have become the primary vehicle for data integration and are critical to the success of modern AI, but the diversity of KG modellin

applicationsarxiv-cs-ai
26 May 2026
Model Releases

QUIET: A Multi-Blank Cascaded Story Cloze Benchmark for LLM Creative Generation Capability

DGX agent

arXiv:2605.25955v1 Announce Type: cross Abstract: Large language models (LLMs) face a dual challenge in creative capability evaluation: existing benchmarks (e.g., Story Cloze Test, HellaSwag) measure

model-releasesarxiv-cs-ai
26 May 2026
Agents

Reward Shaping and Action Masking for Compositional Tasks using Behavior Trees and LLMs

DGX agent

arXiv:2605.05795v2 Announce Type: replace Abstract: Decomposing complex tasks into a sequence of simpler subtasks can improve learning efficiency for an autonomous agent. Reinforcement learning (RL) c

agentsarxiv-cs-lg
26 May 2026
Safety

RiskBridge: Turning CVEs into Business-Aligned Patch Priorities

DGX agent

arXiv:2601.06201v2 Announce Type: replace-cross Abstract: Enterprises are confronted with an unprecedented escalation in cybersecurity vulnerabilities, with thousands of new CVEs disclosed each month.

safetyarxiv-cs-ai
26 May 2026
Agents

A Proactive Multi-Agent Dialogue Framework for Assessing Social Language Disorder Traits in Autism

DGX agent

arXiv:2605.22993v1 Announce Type: cross Abstract: Characteristic linguistic behaviors associated with Social Language Disorder (SLD) in autism spectrum disorder, including echoic repetition, pronoun d

agentsarxiv-cs-ai
25 May 2026
Research

Reinforced Graph of Thoughts: RL-Driven Adaptive Prompting for LLMs

DGX agent

arXiv:2605.22195v1 Announce Type: new Abstract: Graph of Thoughts (GoT), a generalized form of recent prompting paradigms for large language models (LLMs), has been shown to be useful for elaborate pr

researcharxiv-cs-lg
23 May 2026
Model Releases

InternBootcamp Technical Report: Boosting LLM Reasoning with Verifiable Task Scaling

DGX agent

arXiv:2508.08636v2 Announce Type: replace Abstract: Large language models (LLMs) have revolutionized artificial intelligence by enabling complex reasoning capabilities. While recent advancements in re

model-releasesarxiv-cs-cl
21 May 2026
Research

Towards Integrated Rock Support Visualisation in 3D Point Cloud of Underground Mines

DGX agent

arXiv:2605.20973v1 Announce Type: new Abstract: The effectiveness of rock support in underground mines depends on the interaction between installed rock bolts and the structural fabric of the surround

researcharxiv-cs-cv
21 May 2026
Applications

Validating Navmesh using Geometry: Voxel-Based Analysis with Prioritized Exploration

DGX agent

arXiv:2605.21397v1 Announce Type: cross Abstract: Navigation mesh (Navmesh) inconsistencies affect the player experience by directly impacting the navigation systems used by non-playable characters (N

applicationsarxiv-cs-ro
21 May 2026
Model Releases

BLINKG: A Benchmark for LLM-Integrated Knowledge Graph Generation

DGX agent

arXiv:2605.19518v1 Announce Type: new Abstract: Generating Knowledge Graphs (KGs) remains one of the most time-consuming and labor-intensive tasks for knowledge engineers, as they need to identify sem

model-releasesarxiv-cs-ai
20 May 2026
Research

MapAnything: Evaluating Monocular Metric Depth Models for 3D Urban Asset Localization

DGX agent

arXiv:2509.14839v2 Announce Type: replace Abstract: City administrations increasingly rely on comprehensive databases and urban digital twins of city assets, such as traffic signs and trees, as well a

researcharxiv-cs-cv
20 May 2026
Agents

Operationalising Artificial Intelligence Bills of Materials (AIBOMs) for Verifiable AI Provenance and Lifecycle Assurance

DGX agent

arXiv:2605.19755v1 Announce Type: cross Abstract: Artificial Intelligence (AI) systems are increasingly dependent on complex, multi-layered software supply chains that introduce challenges for reprodu

agentsarxiv-cs-ai
20 May 2026
Local Ai

ALIGN: A Vision-Language Framework for High-Accuracy Accident Location Inference through Geo-Spatial Neural Reasoning

DGX agent

arXiv:2511.06316v3 Announce Type: replace Abstract: In low- and middle-income countries, public safety and urban planning initiatives frequently face a critical shortage of accurate, location-specific

local-aiarxiv-cs-ai
19 May 2026
Agents

EnvFactory: Scaling Tool-Use Agents via Executable Environments Synthesis and Robust RL

DGX agent

arXiv:2605.18703v1 Announce Type: new Abstract: Equipping LLMs with tool-use capabilities via Agentic Reinforcement Learning (Agentic RL) is bottlenecked by two challenges: the lack of scalable, robus

agentsarxiv-cs-cl
19 May 2026
Research

Knowledge-to-Verification: Exploring RLVR for LLMs in Knowledge-Intensive Domains

DGX agent

arXiv:2605.18261v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has demonstrated promising potential to enhance the reasoning capabilities of large language model

researcharxiv-cs-cl
19 May 2026
Model Releases

Multilingual jailbreaking of LLMs using low-resource languages

DGX agent

arXiv:2605.18239v1 Announce Type: cross Abstract: Large Language Models (LLMs) remain vulnerable to jailbreak attempts that circumvent safety guardrails. We investigate whether multi-turn conversation

model-releasesarxiv-cs-ai
19 May 2026
Safety

RL4RLA: Teaching ML to Discover Randomized Linear Algebra Algorithms Through Curriculum Design and Graph-Based Search

DGX agent

arXiv:2605.18004v1 Announce Type: new Abstract: Randomized linear algebra (RLA) algorithms are a modern class of numerical linear algebra techniques that play an essential role in scientific computing

safetyarxiv-cs-lg
19 May 2026
Research

Velocity and stroke rate reconstruction of canoe sprint team boats based on panned and zoomed video recordings

DGX agent

arXiv:2602.22941v2 Announce Type: replace Abstract: Pacing strategies, defined by velocity and stroke rate profiles, are essential for peak performance in canoe sprint. While GPS is the gold standard

researcharxiv-cs-cv
19 May 2026
Model Releases

RTL-BenchMT: Dynamic Maintenance of RTL Generation Benchmark Through Agent-Assisted Analysis and Revision

DGX agent

arXiv:2605.15537v1 Announce Type: new Abstract: This paper introduces RTL-BenchMT, an agentic framework for dynamically maintaining RTL generation benchmarks. Large Language Models (LLMs) assisted aut

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Databricks brings GPT-5.5 to enterprise agent workflows

DGX agent

Databricks has integrated OpenAI's GPT-5.5 model into enterprise agent workflows, enabling organizations to build and deploy AI agents with advanced language capabilities. This partnership leverages D

model-releasesopenai
15 May 2026
Tutorials

I strongly believe there are entire companies right now under heavy AI psychosis and its impossible to have rational conversations about it …

DGX agent

I strongly believe there are entire companies right now under heavy AI psychosis and its impossible to have rational conversations about it with them. I can't name any specific people because they inc

tutorialsjeremy-howard--x
15 May 2026
Agents

Lang2MLIP: End-to-End Language-to-Machine Learning Interatomic Potential Development with Autonomous Agentic Workflows

DGX agent

arXiv:2605.14527v1 Announce Type: new Abstract: Developing machine learning interatomic potentials (MLIPs) for complex materials systems remains challenging because it requires expertise in atomistic

agentsarxiv-cs-lg
15 May 2026
Model Releases

OPT-Engine: Benchmarking the Limits of LLMs in Optimization Modeling via Complexity Scaling

DGX agent

arXiv:2601.19924v2 Announce Type: replace-cross Abstract: We investigate the capabilities and scalability of Large Language Models (LLMs) in optimization modeling, a domain requiring structured reason

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

scShapeBench: Discovering geometry from high dimensional scRNAseq data

DGX agent

arXiv:2605.12662v1 Announce Type: new Abstract: High-dimensional point cloud data arise across many scientific domains, especially single-cell biology. The shapes or topologies of these datasets deter

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

FLAME: A New Dataset on FLemish Accounts of Momentary Experiences

DGX agent

arXiv:2504.14707v3 Announce Type: replace Abstract: We introduce FLAME (FLemish Accounts of Momentary Experiences), a new corpus of nearly 25,000 daily personal narratives in Belgian-Dutch (Flemish),

model-releasesarxiv-cs-cl
13 May 2026
Safety

The new era of SaMD: Why cloud infrastructure is the foundation for digital health in 2026

DGX agent

In the healthcare and life sciences industries, speed saves lives, but meeting regulatory requirements and other administrative burdens often pumps the brakes for manufacturers of software as a medica

safetygoogle-cloud-ai
13 May 2026
← Previous
1…4849505152…104
Next →