AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “automated”

GridTimelineEvolution
4,952 results
28 Apr 2026

xOffense: An Autonomous Multi-Agent Framework for Penetration Testing with Domain-Adapted Large Language Models

Model ReleasesDGX agent

arXiv:2509.13021v2 Announce Type: replace-cross Abstract: This work introduces xOffense, an AI-driven, multi-agent penetration testing framework that shifts the process from labor-intensive, expert-dr

Your Students Don't Use LLMs Like You Wish They Did

SafetyDGX agent

arXiv:2604.23486v1 Announce Type: new Abstract: Educational NLP systems are typically evaluated using engagement metrics and satisfaction surveys, which are at best a proxy for meeting pedagogical goa

ZenBrain: A Neuroscience-Inspired 7-Layer Memory Architecture for Autonomous AI Systems

SafetyDGX agent

arXiv:2604.23878v1 Announce Type: new Abstract: Despite a century of empirical memory research, existing AI agent memory systems rely on system-engineering metaphors (virtual-memory paging, flat LLM s


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
27 Apr 2026

Atlas-Alignment: Making Interpretability Transferable Across Language Models

Model ReleasesDGX agent

arXiv:2510.27413v2 Announce Type: replace-cross Abstract: Interpretability is crucial for building safe, reliable, and controllable language models, yet existing interpretability pipelines remain cost

Can Large Language Models Adequately Perform Symbolic Reasoning Over Time Series?

Model ReleasesDGX agent

arXiv:2508.03963v4 Announce Type: replace Abstract: Uncovering hidden symbolic laws from time series data, as an aspiration dating back to Kepler's discovery of planetary motion, remains a core challe

ChangeQuery: Advancing Remote Sensing Change Analysis for Natural and Human-Induced Disasters from Visual Detection to Semantic Understanding

Model ReleasesDGX agent

arXiv:2604.22333v1 Announce Type: cross Abstract: Rapid situational awareness is critical in post-disaster response. While remote sensing damage assessment is evolving from pixel-level change detectio

CRAFT: Clustered Regression for Adaptive Filtering of Training data

ResearchDGX agent

arXiv:2604.22693v1 Announce Type: cross Abstract: Selecting a small, high-quality subset from a large corpus for fine-tuning is increasingly important as corpora grow to tens of millions of datapoints

Emergent Strategic Reasoning Risks in AI: A Taxonomy-Driven Evaluation Framework

SafetyDGX agent

arXiv:2604.22119v1 Announce Type: new Abstract: As reasoning capacity and deployment scope grow in tandem, large language models (LLMs) gain the capacity to engage in behaviors that serve their own ob

FeatEHR-LLM: Leveraging Large Language Models for Feature Engineering in Electronic Health Records

ApplicationsDGX agent

arXiv:2604.22534v1 Announce Type: cross Abstract: Feature engineering for Electronic Health Records (EHR) is complicated by irregular observation intervals, variable measurement frequencies, and struc

FixV2W: Correcting Invalid CVE-CWE Mappings with Knowledge Graph Embeddings

ResearchDGX agent

arXiv:2604.22176v1 Announce Type: cross Abstract: Accurate mapping between Common Vulnerabilities and Exposures (CVE) and Common Weakness Enumeration (CWE) entries is critical for effective vulnerabil

Learning Evidence Highlighting for Frozen LLMs

SafetyDGX agent

arXiv:2604.22565v1 Announce Type: cross Abstract: Large Language Models (LLMs) can reason well, yet often miss decisive evidence when it is buried in long, noisy contexts. We introduce HiLight, an Evi

Lifting Unlabeled Internet-level Data for 3D Scene Understanding

ResearchDGX agent

arXiv:2604.01907v2 Announce Type: replace-cross Abstract: Annotated 3D scene data is scarce and expensive to acquire, while abundant unlabeled videos are readily available on the internet. In this pap

Manifold Learning for Personalized and Label-Free Detection of Cardiac Arrhythmias

ApplicationsDGX agent

arXiv:2506.16494v3 Announce Type: replace Abstract: Electrocardiograms (ECGs) provide non-invasive measurements of heart activity and are established tools for detecting cardiac arrhythmias. Although

Memanto: Typed Semantic Memory with Information-Theoretic Retrieval for Long-Horizon Agents

AgentsDGX agent

arXiv:2604.22085v1 Announce Type: new Abstract: The transition from stateless language model inference to persistent, multi session autonomous agents has revealed memory to be a primary architectural

Predicting Liquidity-Aware Bond Yields using Causal GANs and Deep Reinforcement Learning with LLM Evaluation

ApplicationsDGX agent

arXiv:2502.17011v2 Announce Type: replace-cross Abstract: Financial bond yield forecasting is challenging due to data scarcity, nonlinear macroeconomic dependencies, and evolving market conditions. In

PrivSTRUCT: Untangling Data Purpose Compliance of Privacy Policies in Google Play Store

TutorialsDGX agent

arXiv:2604.22157v1 Announce Type: cross Abstract: Existing research typically treats privacy policies as flat, uniform text, extracting information without regard for the document's logical hierarchy.

Read the release notes: https://github.com/keras-team/kinetic/releases/tag/0.0.2

ResearchDGX agent

Kinetic version 0.0.2 release notes are available on GitHub, detailing updates and features for this Keras-related project. The announcement was shared by François Chollet, Keras creator, on X (former

Rethinking XAI Evaluation: A Human-Centered Audit of Shapley Benchmarks in High-Stakes Settings

SafetyDGX agent

arXiv:2604.22662v1 Announce Type: cross Abstract: Shapley values are a cornerstone of explainable AI, yet their proliferation into competing formulations has created a fragmented landscape with little

Robust Localization for Autonomous Vehicles in Highway Scenes

AgentsDGX agent

arXiv:2604.22040v1 Announce Type: new Abstract: Localization for autonomous vehicles on highways remains under-explored compared to urban roads, and state-of-the-art methods for urban scenes degrade w

The Biggest Risk of Embodied AI is Governance Lag

SafetyDGX agent

arXiv:2604.21938v1 Announce Type: cross Abstract: Embodied AI is widely discussed as a job-displacement problem. The deeper risk, however, is governance lag: the inability of public institutions to ke

TRACE: Topology-aware Reconstruction of Accidents in CARLA for AV Evaluation

Model ReleasesDGX agent

arXiv:2604.22068v1 Announce Type: cross Abstract: Validating Autonomous Vehicles (AVs) requires exposure to rare, safety-critical scenarios, infrequent in routine driving data. Existing benchmarks add

26 Apr 2026

Epoch AI: Google controls ~25% of global AI compute, with ~3.8M TPUs and 1.3M GPUs; Google Cloud CEO Thomas Kurian says demand and revenue justify the spend (Stephen Morris/Financial Times)

IndustryDGX agent

Stephen Morris / Financial Times: Epoch AI: Google controls ~25% of global AI compute, with ~3.8M TPUs and 1.3M GPUs; Google Cloud CEO Thomas Kurian says demand and revenue justify the spend — Thomas

25 Apr 2026

gpt-5.5 is now available in the ml-intern! this means it gets access to the whole @huggingface infra: buckets, jobs, repos etc for doing ai …

Model ReleasesDGX agent

gpt-5.5 is now available in the ml-intern! this means it gets access to the whole @huggingface infra: buckets, jobs, repos etc for doing ai research at scale giving it a spin now to see if i'm even cl

24 Apr 2026

AI for software engineering: from probable to provable

ResearchDGX agent

arXiv:2511.23159v2 Announce Type: replace-cross Abstract: Vibe coding, the much-touted use of AI techniques for programming, faces two overwhelming obstacles: the difficulty of specifying goals ('prom

AI Governance under Political Turnover: The Alignment Surface of Compliance Design

SafetyDGX agent

arXiv:2604.21103v1 Announce Type: new Abstract: Governments are increasingly interested in using AI to make administrative decisions cheaper, more scalable, and more consistent. But for probabilistic

Autobots, assemble!

IndustryDGX agent

Elon Musk posted about Tesla's Optimus humanoid robot, likely announcing a development milestone or demonstrating capabilities of the autonomous robot project. The post uses the 'Autobots, assemble!'

Conjecture and Inquiry: Quantifying Software Performance Requirements via Interactive Retrieval-Augmented Preference Elicitation

ApplicationsDGX agent

arXiv:2604.21380v1 Announce Type: cross Abstract: Since software performance requirements are documented in natural language, quantifying them into mathematical forms is essential for software enginee

Deep FinResearch Bench: Evaluating AI's Ability to Conduct Professional Financial Investment Research

Model ReleasesDGX agent

arXiv:2604.21006v1 Announce Type: new Abstract: We introduce Deep FinResearch Bench, a practical and comprehensive evaluation framework for deep research (DR) agents in financial investment research.

DiagramBank: A Large-scale Dataset of Diagram Design Exemplars with Paper Metadata for Retrieval-Augmented Generation

AgentsDGX agent

arXiv:2604.20857v1 Announce Type: cross Abstract: Recent advances in autonomous ``AI scientist'' systems have demonstrated the ability to automatically write scientific manuscripts and codes with exec

Doubly Saturated Ramsey Graphs: A Case Study in Computer-Assisted Mathematical Discovery

ApplicationsDGX agent

arXiv:2604.21187v1 Announce Type: cross Abstract: Ramsey-good graphs are graphs that contain neither a clique of size s nor an independent set of size t. We study doubly saturated Ramsey-good graphs,

EduCoder: An Open-Source Annotation System for Education Transcript Data

ApplicationsDGX agent

arXiv:2507.05385v4 Announce Type: replace Abstract: We introduce EduCoder, a domain-specialized tool designed to support utterance-level annotation of educational dialogue. While general-purpose text

Enhancing Online Recruitment with Category-Aware MoE and LLM-based Data Augmentation

ResearchDGX agent

arXiv:2604.21264v1 Announce Type: new Abstract: Person-Job Fit (PJF) is a critical component for online recruitment. Existing approaches face several challenges, particularly in handling low-quality j

Enhancing Science Classroom Discourse Analysis through Joint Multi-Task Learning for Reasoning-Component Classification

Model ReleasesDGX agent

arXiv:2604.21137v1 Announce Type: cross Abstract: Analyzing the reasoning patterns of students in science classrooms is critical for understanding knowledge construction mechanism and improving instru

Escaping the Agreement Trap: Defensibility Signals for Evaluating Rule-Governed AI

SafetyDGX agent

arXiv:2604.20972v1 Announce Type: new Abstract: Content moderation systems are typically evaluated by measuring agreement with human labels. In rule-governed environments this assumption fails: multip

EVENT5Ws: A Large Dataset for Open-Domain Event Extraction from Documents

Model ReleasesDGX agent

arXiv:2604.21890v1 Announce Type: new Abstract: Event extraction identifies the central aspects of events from text. It supports event understanding and analysis, which is crucial for tasks such as in

For the last 72 hours since ml-intern launched we have had over 500+ autonomous AI research projects running on the Space at all times. Some…

Model ReleasesDGX agent

For the last 72 hours since ml-intern launched we have had over 500+ autonomous AI research projects running on the Space at all times. Some insane ones I saw: 1. A new AI paradigm from scratch — tryi

FunduSegmenter: Leveraging the RETFound Foundation Model for Joint Optic Disc and Optic Cup Segmentation in Retinal Fundus Images

ResearchDGX agent

arXiv:2508.11354v3 Announce Type: replace-cross Abstract: Purpose: This study introduces the first adaptation of RETFound for joint optic disc (OD) and optic cup (OC) segmentation. RETFound is a well-

Grounding Machine Creativity in Game Design Knowledge Representations: Empirical Probing of LLM-Based Executable Synthesis of Goal Playable Patterns under Structural Constraints

Model ReleasesDGX agent

arXiv:2603.07101v3 Announce Type: replace Abstract: Creatively translating complex gameplay ideas into executable artifacts (e.g., games as Unity projects and code) remains a central challenge in comp

HWE-Bench: Benchmarking LLM Agents on Real-World Hardware Bug Repair Tasks

Model ReleasesDGX agent

arXiv:2604.14709v2 Announce Type: replace Abstract: Existing benchmarks for hardware design primarily evaluate Large Language Models (LLMs) on isolated, component-level tasks such as generating HDL mo

Integrated packing, placement, scheduling, and routing of personalized production: a pharmaceutical Industry 4.0 use-case with a planar transport system

ApplicationsDGX agent

arXiv:2604.21029v1 Announce Type: cross Abstract: The recent emergence of planar transport systems necessitates re-evaluation of Flexible Manufacturing Systems (FMS) to address the simultaneous schedu

Multimodal Bayesian Network for Robust Assessment of Casualties in Autonomous Triage

AgentsDGX agent

arXiv:2512.18908v2 Announce Type: replace Abstract: Mass Casualty Incidents can overwhelm emergency medical systems and resulting delays or errors in the assessment of casualties can lead to preventab

PLAS-Net: Pixel-Level Area Segmentation for UAV-Based Beach Litter Monitoring

ResearchDGX agent

arXiv:2604.21313v1 Announce Type: new Abstract: Accurate quantification of the physical exposure area of beach litter, rather than simple item counts, is essential for credible ecological risk assessm

Really impressed by how smooth switching most of my coding tasks to Codex (GPT-5.5) from Claude Code (Opus 4.7) has been. I thought it was g…

Model ReleasesDGX agent

Really impressed by how smooth switching most of my coding tasks to Codex (GPT-5.5) from Claude Code (Opus 4.7) has been. I thought it was going to be more difficult and that I would be 'fighting' wit

Scensory: Real-Time Robotic Olfactory Perception for Joint Identification and Source Localization

ResearchDGX agent

arXiv:2509.19318v2 Announce Type: replace-cross Abstract: While robotic perception has advanced rapidly in vision and touch, enabling robots to reason about indoor fungal contamination from weak, diff

SemEval-2026 Task 4: Narrative Story Similarity and Narrative Representation Learning

ResearchDGX agent

arXiv:2604.21782v1 Announce Type: new Abstract: We present the shared task on narrative similarity and narrative representation learning - NSNRL (pronounced 'nass-na-rel'). The task operationalizes na

SocraticKG: Knowledge Graph Construction via QA-Driven Fact Extraction

Model ReleasesDGX agent

arXiv:2601.10003v2 Announce Type: replace Abstract: Constructing Knowledge Graphs (KGs) from unstructured text provides a structured framework for knowledge representation and reasoning, yet current L

Structural Quality Gaps in Practitioner AI Governance Prompts: An Empirical Study Using a Five-Principle Evaluation Framework

AgentsDGX agent

arXiv:2604.21090v1 Announce Type: cross Abstract: AI governance programmes increasingly rely on natural language prompts to constrain and direct AI agent behaviour. These prompts function as executabl

The Economics of p(doom): Scenarios of Existential Risk and Economic Growth in the Age of Transformative AI

SafetyDGX agent

arXiv:2503.07341v2 Announce Type: replace-cross Abstract: Recent advances in artificial intelligence (AI) have led to a wide range of predictions about its long-term impact on humanity. A central focu

VG-CoT: Towards Trustworthy Visual Reasoning via Grounded Chain-of-Thought

Model ReleasesDGX agent

arXiv:2604.21396v1 Announce Type: cross Abstract: The advancement of Large Vision-Language Models (LVLMs) requires precise local region-based reasoning that faithfully grounds the model's logic in act

XtraGPT: Context-Aware and Controllable Academic Paper Revision via Human-AI Collaboration

SafetyDGX agent

arXiv:2505.11336v4 Announce Type: replace Abstract: Despite the growing adoption of large language models (LLMs) in academic workflows, their capabilities remain limited in supporting high-quality sci

23 Apr 2026

Automatic Ontology Construction Using LLMs as an External Layer of Memory, Verification, and Planning for Hybrid Intelligent Systems

Model ReleasesDGX agent

arXiv:2604.20795v1 Announce Type: new Abstract: This paper presents a hybrid architecture for intelligent systems in which large language models (LLMs) are extended with an external ontological memory

Beyond the Crowd: LLM-Augmented Community Notes for Governing Health Misinformation

Model ReleasesDGX agent

arXiv:2510.11423v3 Announce Type: replace-cross Abstract: Community Notes, the crowd-sourced misinformation governance system on X (formerly Twitter), allows users to flag misleading posts, attach con

Bias in the Tails: How Name-conditioned Evaluative Framing in Resume Summaries Destabilizes LLM-based Hiring

SafetyDGX agent

arXiv:2604.19984v1 Announce Type: cross Abstract: Research has documented LLMs' name-based bias in hiring and salary recommendations. In this paper, we instead consider a setting where LLMs generate c

ChipCraftBrain: Validation-First RTL Generation via Multi-Agent Orchestration

SafetyDGX agent

arXiv:2604.19856v1 Announce Type: cross Abstract: Large Language Models (LLMs) show promise for generating Register-Transfer Level (RTL) code from natural language specifications, but single-shot gene

Co-Located Tests, Better AI Code: How Test Syntax Structure Affects Foundation Model Code Generation

ResearchDGX agent

arXiv:2604.19826v1 Announce Type: cross Abstract: AI coding assistants increasingly generate code alongside tests. How developers structure test code, whether inline with the implementation or in sepa

CXR-LanIC: Language-Grounded Interpretable Classifier for Chest X-Ray Diagnosis

ResearchDGX agent

arXiv:2510.21464v2 Announce Type: replace Abstract: Deep learning models have achieved remarkable accuracy in chest X-ray diagnosis, yet their widespread clinical adoption remains limited by the black

Depression Risk Assessment in Social Media via Large Language Models

ResearchDGX agent

arXiv:2604.19887v1 Announce Type: cross Abstract: Depression is one of the most prevalent and debilitating mental health conditions worldwide, frequently underdiagnosed and undertreated. The prolifera

Development and Preliminary Evaluation of a Domain-Specific Large Language Model for Tuberculosis Care in South Africa

Model ReleasesDGX agent

arXiv:2604.19776v1 Announce Type: new Abstract: Tuberculosis (TB) is one of the world's deadliest infectious diseases, and in South Africa, it contributes a significant burden to the country's health

ESGLens: An LLM-Based RAG Framework for Interactive ESG Report Analysis and Score Prediction

ResearchDGX agent

arXiv:2604.19779v1 Announce Type: new Abstract: Environmental, Social, and Governance (ESG) reports are central to investment decision-making, yet their length, heterogeneous content, and lack of stan

Evals ~= Environments…they’re one of the best investments a team can make for improving agents Step 0: Turn On Tracing for Agents Step 1: Po…

AgentsDGX agent

Evals ~= Environments…they’re one of the best investments a team can make for improving agents Step 0: Turn On Tracing for Agents Step 1: Point compute at Traces to understand agent behavior, segment

← Previous
1…7374757677…83
Next →