AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlog
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “automated”

GridTimelineEvolution
4,952 results
Safety

AI Governance under Political Turnover: The Alignment Surface of Compliance Design

DGX agent

arXiv:2604.21103v1 Announce Type: new Abstract: Governments are increasingly interested in using AI to make administrative decisions cheaper, more scalable, and more consistent. But for probabilistic

safetyarxiv-cs-ai
24 Apr 2026
Industry

Autobots, assemble!

DGX agent
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Elon Musk posted about Tesla's Optimus humanoid robot, likely announcing a development milestone or demonstrating capabilities of the autonomous robot project. The post uses the 'Autobots, assemble!'

industryelon-musk--x
24 Apr 2026
Applications

Conjecture and Inquiry: Quantifying Software Performance Requirements via Interactive Retrieval-Augmented Preference Elicitation

DGX agent

arXiv:2604.21380v1 Announce Type: cross Abstract: Since software performance requirements are documented in natural language, quantifying them into mathematical forms is essential for software enginee

applicationsarxiv-cs-ai
24 Apr 2026
Model Releases

Deep FinResearch Bench: Evaluating AI's Ability to Conduct Professional Financial Investment Research

DGX agent

arXiv:2604.21006v1 Announce Type: new Abstract: We introduce Deep FinResearch Bench, a practical and comprehensive evaluation framework for deep research (DR) agents in financial investment research.

model-releasesarxiv-cs-ai
24 Apr 2026
Agents

DiagramBank: A Large-scale Dataset of Diagram Design Exemplars with Paper Metadata for Retrieval-Augmented Generation

DGX agent

arXiv:2604.20857v1 Announce Type: cross Abstract: Recent advances in autonomous ``AI scientist'' systems have demonstrated the ability to automatically write scientific manuscripts and codes with exec

agentsarxiv-cs-ai
24 Apr 2026
Applications

Doubly Saturated Ramsey Graphs: A Case Study in Computer-Assisted Mathematical Discovery

DGX agent

arXiv:2604.21187v1 Announce Type: cross Abstract: Ramsey-good graphs are graphs that contain neither a clique of size s nor an independent set of size t. We study doubly saturated Ramsey-good graphs,

applicationsarxiv-cs-ai
24 Apr 2026
Applications

EduCoder: An Open-Source Annotation System for Education Transcript Data

DGX agent

arXiv:2507.05385v4 Announce Type: replace Abstract: We introduce EduCoder, a domain-specialized tool designed to support utterance-level annotation of educational dialogue. While general-purpose text

applicationsarxiv-cs-cl
24 Apr 2026
Research

Enhancing Online Recruitment with Category-Aware MoE and LLM-based Data Augmentation

DGX agent

arXiv:2604.21264v1 Announce Type: new Abstract: Person-Job Fit (PJF) is a critical component for online recruitment. Existing approaches face several challenges, particularly in handling low-quality j

researcharxiv-cs-ai
24 Apr 2026
Model Releases

Enhancing Science Classroom Discourse Analysis through Joint Multi-Task Learning for Reasoning-Component Classification

DGX agent

arXiv:2604.21137v1 Announce Type: cross Abstract: Analyzing the reasoning patterns of students in science classrooms is critical for understanding knowledge construction mechanism and improving instru

model-releasesarxiv-cs-ai
24 Apr 2026
Safety

Escaping the Agreement Trap: Defensibility Signals for Evaluating Rule-Governed AI

DGX agent

arXiv:2604.20972v1 Announce Type: new Abstract: Content moderation systems are typically evaluated by measuring agreement with human labels. In rule-governed environments this assumption fails: multip

safetyarxiv-cs-ai
24 Apr 2026
Model Releases

EVENT5Ws: A Large Dataset for Open-Domain Event Extraction from Documents

DGX agent

arXiv:2604.21890v1 Announce Type: new Abstract: Event extraction identifies the central aspects of events from text. It supports event understanding and analysis, which is crucial for tasks such as in

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

For the last 72 hours since ml-intern launched we have had over 500+ autonomous AI research projects running on the Space at all times. Some…

DGX agent

For the last 72 hours since ml-intern launched we have had over 500+ autonomous AI research projects running on the Space at all times. Some insane ones I saw: 1. A new AI paradigm from scratch — tryi

model-releasesclem-delangue--x
24 Apr 2026
Research

FunduSegmenter: Leveraging the RETFound Foundation Model for Joint Optic Disc and Optic Cup Segmentation in Retinal Fundus Images

DGX agent

arXiv:2508.11354v3 Announce Type: replace-cross Abstract: Purpose: This study introduces the first adaptation of RETFound for joint optic disc (OD) and optic cup (OC) segmentation. RETFound is a well-

researcharxiv-cs-ai
24 Apr 2026
Model Releases

Grounding Machine Creativity in Game Design Knowledge Representations: Empirical Probing of LLM-Based Executable Synthesis of Goal Playable Patterns under Structural Constraints

DGX agent

arXiv:2603.07101v3 Announce Type: replace Abstract: Creatively translating complex gameplay ideas into executable artifacts (e.g., games as Unity projects and code) remains a central challenge in comp

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

HWE-Bench: Benchmarking LLM Agents on Real-World Hardware Bug Repair Tasks

DGX agent

arXiv:2604.14709v2 Announce Type: replace Abstract: Existing benchmarks for hardware design primarily evaluate Large Language Models (LLMs) on isolated, component-level tasks such as generating HDL mo

model-releasesarxiv-cs-ai
24 Apr 2026
Applications

Integrated packing, placement, scheduling, and routing of personalized production: a pharmaceutical Industry 4.0 use-case with a planar transport system

DGX agent

arXiv:2604.21029v1 Announce Type: cross Abstract: The recent emergence of planar transport systems necessitates re-evaluation of Flexible Manufacturing Systems (FMS) to address the simultaneous schedu

applicationsarxiv-cs-ai
24 Apr 2026
Agents

Multimodal Bayesian Network for Robust Assessment of Casualties in Autonomous Triage

DGX agent

arXiv:2512.18908v2 Announce Type: replace Abstract: Mass Casualty Incidents can overwhelm emergency medical systems and resulting delays or errors in the assessment of casualties can lead to preventab

agentsarxiv-cs-ai
24 Apr 2026
Research

PLAS-Net: Pixel-Level Area Segmentation for UAV-Based Beach Litter Monitoring

DGX agent

arXiv:2604.21313v1 Announce Type: new Abstract: Accurate quantification of the physical exposure area of beach litter, rather than simple item counts, is essential for credible ecological risk assessm

researcharxiv-cs-cv
24 Apr 2026
Model Releases

Really impressed by how smooth switching most of my coding tasks to Codex (GPT-5.5) from Claude Code (Opus 4.7) has been. I thought it was g…

DGX agent

Really impressed by how smooth switching most of my coding tasks to Codex (GPT-5.5) from Claude Code (Opus 4.7) has been. I thought it was going to be more difficult and that I would be 'fighting' wit

model-releasesdair-ai--x
24 Apr 2026
Research

Scensory: Real-Time Robotic Olfactory Perception for Joint Identification and Source Localization

DGX agent

arXiv:2509.19318v2 Announce Type: replace-cross Abstract: While robotic perception has advanced rapidly in vision and touch, enabling robots to reason about indoor fungal contamination from weak, diff

researcharxiv-cs-ro
24 Apr 2026
Research

SemEval-2026 Task 4: Narrative Story Similarity and Narrative Representation Learning

DGX agent

arXiv:2604.21782v1 Announce Type: new Abstract: We present the shared task on narrative similarity and narrative representation learning - NSNRL (pronounced 'nass-na-rel'). The task operationalizes na

researcharxiv-cs-cl
24 Apr 2026
Model Releases

SocraticKG: Knowledge Graph Construction via QA-Driven Fact Extraction

DGX agent

arXiv:2601.10003v2 Announce Type: replace Abstract: Constructing Knowledge Graphs (KGs) from unstructured text provides a structured framework for knowledge representation and reasoning, yet current L

model-releasesarxiv-cs-cl
24 Apr 2026
Agents

Structural Quality Gaps in Practitioner AI Governance Prompts: An Empirical Study Using a Five-Principle Evaluation Framework

DGX agent

arXiv:2604.21090v1 Announce Type: cross Abstract: AI governance programmes increasingly rely on natural language prompts to constrain and direct AI agent behaviour. These prompts function as executabl

agentsarxiv-cs-ai
24 Apr 2026
Safety

The Economics of p(doom): Scenarios of Existential Risk and Economic Growth in the Age of Transformative AI

DGX agent

arXiv:2503.07341v2 Announce Type: replace-cross Abstract: Recent advances in artificial intelligence (AI) have led to a wide range of predictions about its long-term impact on humanity. A central focu

safetyarxiv-cs-ai
24 Apr 2026
Model Releases

VG-CoT: Towards Trustworthy Visual Reasoning via Grounded Chain-of-Thought

DGX agent

arXiv:2604.21396v1 Announce Type: cross Abstract: The advancement of Large Vision-Language Models (LVLMs) requires precise local region-based reasoning that faithfully grounds the model's logic in act

model-releasesarxiv-cs-ai
24 Apr 2026
Safety

XtraGPT: Context-Aware and Controllable Academic Paper Revision via Human-AI Collaboration

DGX agent

arXiv:2505.11336v4 Announce Type: replace Abstract: Despite the growing adoption of large language models (LLMs) in academic workflows, their capabilities remain limited in supporting high-quality sci

safetyarxiv-cs-cl
24 Apr 2026
Model Releases

Automatic Ontology Construction Using LLMs as an External Layer of Memory, Verification, and Planning for Hybrid Intelligent Systems

DGX agent

arXiv:2604.20795v1 Announce Type: new Abstract: This paper presents a hybrid architecture for intelligent systems in which large language models (LLMs) are extended with an external ontological memory

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Beyond the Crowd: LLM-Augmented Community Notes for Governing Health Misinformation

DGX agent

arXiv:2510.11423v3 Announce Type: replace-cross Abstract: Community Notes, the crowd-sourced misinformation governance system on X (formerly Twitter), allows users to flag misleading posts, attach con

model-releasesarxiv-cs-cl
23 Apr 2026
Safety

Bias in the Tails: How Name-conditioned Evaluative Framing in Resume Summaries Destabilizes LLM-based Hiring

DGX agent

arXiv:2604.19984v1 Announce Type: cross Abstract: Research has documented LLMs' name-based bias in hiring and salary recommendations. In this paper, we instead consider a setting where LLMs generate c

safetyarxiv-cs-ai
23 Apr 2026
Safety

ChipCraftBrain: Validation-First RTL Generation via Multi-Agent Orchestration

DGX agent

arXiv:2604.19856v1 Announce Type: cross Abstract: Large Language Models (LLMs) show promise for generating Register-Transfer Level (RTL) code from natural language specifications, but single-shot gene

safetyarxiv-cs-ai
23 Apr 2026
Research

Co-Located Tests, Better AI Code: How Test Syntax Structure Affects Foundation Model Code Generation

DGX agent

arXiv:2604.19826v1 Announce Type: cross Abstract: AI coding assistants increasingly generate code alongside tests. How developers structure test code, whether inline with the implementation or in sepa

researcharxiv-cs-ai
23 Apr 2026
Research

CXR-LanIC: Language-Grounded Interpretable Classifier for Chest X-Ray Diagnosis

DGX agent

arXiv:2510.21464v2 Announce Type: replace Abstract: Deep learning models have achieved remarkable accuracy in chest X-ray diagnosis, yet their widespread clinical adoption remains limited by the black

researcharxiv-cs-cv
23 Apr 2026
Research

Depression Risk Assessment in Social Media via Large Language Models

DGX agent

arXiv:2604.19887v1 Announce Type: cross Abstract: Depression is one of the most prevalent and debilitating mental health conditions worldwide, frequently underdiagnosed and undertreated. The prolifera

researcharxiv-cs-ai
23 Apr 2026
Model Releases

Development and Preliminary Evaluation of a Domain-Specific Large Language Model for Tuberculosis Care in South Africa

DGX agent

arXiv:2604.19776v1 Announce Type: new Abstract: Tuberculosis (TB) is one of the world's deadliest infectious diseases, and in South Africa, it contributes a significant burden to the country's health

model-releasesarxiv-cs-cl
23 Apr 2026
Research

ESGLens: An LLM-Based RAG Framework for Interactive ESG Report Analysis and Score Prediction

DGX agent

arXiv:2604.19779v1 Announce Type: new Abstract: Environmental, Social, and Governance (ESG) reports are central to investment decision-making, yet their length, heterogeneous content, and lack of stan

researcharxiv-cs-cl
23 Apr 2026
Agents

Evals ~= Environments…they’re one of the best investments a team can make for improving agents Step 0: Turn On Tracing for Agents Step 1: Po…

DGX agent

Evals ~= Environments…they’re one of the best investments a team can make for improving agents Step 0: Turn On Tracing for Agents Step 1: Point compute at Traces to understand agent behavior, segment

agentsharrison-chase--x
23 Apr 2026
Model Releases

Evian: Towards Explainable Visual Instruction-tuning Data Auditing

DGX agent

arXiv:2604.20544v1 Announce Type: cross Abstract: The efficacy of Large Vision-Language Models (LVLMs) is critically dependent on the quality of their training data, requiring a precise balance betwee

model-releasesarxiv-cs-ai
23 Apr 2026
Agents

EvolveSignal: A Large Language Model Powered Coding Agent for Discovering Traffic Signal Control Strategies

DGX agent

arXiv:2509.03335v3 Announce Type: replace Abstract: In traffic engineering, fixed-time traffic signal control remains widely used for its low cost, stability, and interpretability. However, its design

agentsarxiv-cs-lg
23 Apr 2026
Model Releases

Exploring Spatial Intelligence from a Generative Perspective

DGX agent

arXiv:2604.20570v1 Announce Type: new Abstract: Spatial intelligence is essential for multimodal large language models, yet current benchmarks largely assess it only from an understanding perspective.

model-releasesarxiv-cs-cv
23 Apr 2026
Model Releases

From Recall to Forgetting: Benchmarking Long-Term Memory for Personalized Agents

DGX agent

arXiv:2604.20006v1 Announce Type: new Abstract: Personalized agents that interact with users over long periods must maintain persistent memory across sessions and update it as circumstances change. Ho

model-releasesarxiv-cs-cl
23 Apr 2026
Safety

Generative Augmentation of Imbalanced Flight Records for Flight Diversion Prediction: A Multi-objective Optimisation Framework

DGX agent

arXiv:2604.20288v1 Announce Type: new Abstract: Flight diversions are rare but high-impact events in aviation, making their reliable prediction vital for both safety and operational efficiency. Howeve

safetyarxiv-cs-lg
23 Apr 2026
Research

HumorRank: A Tournament-Based Leaderboard for Evaluating Humor Generation in Large Language Models

DGX agent

arXiv:2604.19786v1 Announce Type: new Abstract: Evaluating humor in large language models (LLMs) is an open challenge because existing approaches yield isolated, incomparable metrics rather than unifi

researcharxiv-cs-cl
23 Apr 2026
Applications

If you're interested in production doc parsing, come check out LlamaParse: https://cloud.llamaindex.ai/

DGX agent

LlamaParse is a document parsing tool offered by LlamaIndex that specializes in extracting and processing information from production documents. The tool appears designed to handle complex document fo

applicationsjerry-liu--x
23 Apr 2026
Model Releases

IMPACT-CYCLE: A Contract-Based Multi-Agent System for Claim-Level Supervisory Correction of Long-Video Semantic Memory

DGX agent

arXiv:2604.20136v1 Announce Type: cross Abstract: Correcting errors in long-video understanding is disproportionately costly: existing multimodal pipelines produce opaque, end-to-end outputs that expo

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Learning to Evolve: A Self-Improving Framework for Multi-Agent Systems via Textual Parameter Graph Optimization

DGX agent

arXiv:2604.20714v1 Announce Type: new Abstract: Designing and optimizing multi-agent systems (MAS) is a complex, labor-intensive process of 'Agent Engineering.' Existing automatic optimization methods

model-releasesarxiv-cs-ai
23 Apr 2026
Safety

More people will die from suppressing AI than from the imaginary AI apocalypse. They'll die from restricting safe self-driving cars that are…

DGX agent

More people will die from suppressing AI than from the imaginary AI apocalypse. They'll die from restricting safe self-driving cars that are 90% better drivers than people who kill 1.5 million people

safetyyann-lecun--x
23 Apr 2026
Agents

Open-source agent for long-horizon deep research https://github.com/TIGER-AI-Lab/OpenResearcher

DGX agent

OpenResearcher is an open-source AI agent designed to conduct long-horizon deep research tasks, enabling autonomous investigation and analysis across extended research workflows. The project is mainta

agentsclem-delangue--x
23 Apr 2026
Model Releases

🚨 OpenAI just launched GPT-5.5. The OpenAI team was nice enough to give me early access over the last several weeks, and I just want to fla…

DGX agent

🚨 OpenAI just launched GPT-5.5. The OpenAI team was nice enough to give me early access over the last several weeks, and I just want to flag: there is a certain class of models (one that we’re hitting

model-releasesallie-k--miller--x
23 Apr 2026
← Previous
1…9293949596…104
Next →