AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “automated”

GridTimelineEvolution
4,983 results
10 Apr 2026

LifeAlign: Lifelong Alignment for Large Language Models with Memory-Augmented Focalized Preference Optimization

Model ReleasesDGX agent

arXiv:2509.17183v3 Announce Type: replace-cross Abstract: Alignment plays a crucial role in Large Language Models (LLMs) in aligning with human preferences on a specific task/domain. Traditional align

MemReader: From Passive to Active Extraction for Long-Term Agent Memory

SafetyDGX agent

arXiv:2604.07877v1 Announce Type: new Abstract: Long-term memory is fundamental for personalized and autonomous agents, yet populating it remains a bottleneck. Existing systems treat memory extraction

Mitigating Domain Drift in Multi Species Segmentation with DINOv2: A Cross-Domain Evaluation in Herbicide Research Trials

ApplicationsDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2508.07514v3 Announce Type: replace Abstract: Reliable plant species and damage segmentation for herbicide field research trials requires models that can withstand substantial real-world variati

Near-100% Accurate Data for your Agent with Comprehensive Context Engineering

Model ReleasesDGX agent

Agentic workflows are already used for initiating action. To be successful, agents typically need to combine multiple steps and execute business logic reflective of real-life decisions. But, as develo

Phantom: Physics-Infused Video Generation via Joint Modeling of Visual and Latent Physical Dynamics

ApplicationsDGX agent

arXiv:2604.08503v1 Announce Type: new Abstract: Recent advances in generative video modeling, driven by large-scale datasets and powerful architectures, have yielded remarkable visual realism. However

QueryData helps agents turn natural language into queries for AlloyDB, Cloud SQL and Spanner

Model ReleasesDGX agent

QueryData launches in preview today. It is a tool for translating natural language into database queries with near-100% accuracy. With QueryData, you can build agentic experiences across AlloyDB, Clou

REVEAL: Reasoning-Enhanced Forensic Evidence Analysis for Explainable AI-Generated Image Detection

Model ReleasesDGX agent

arXiv:2511.23158v2 Announce Type: replace-cross Abstract: The rapid progress of visual generative models has made AI-generated images increasingly difficult to distinguish from authentic ones, posing

Riemann-Bench: A Benchmark for Moonshot Mathematics

Model ReleasesDGX agent

arXiv:2604.06802v1 Announce Type: new Abstract: Recent AI systems have achieved gold-medal-level performance on the International Mathematical Olympiad, demonstrating remarkable proficiency at competi

Say Something Else: Rethinking Contextual Privacy as Information Sufficiency

ResearchDGX agent

arXiv:2604.06409v1 Announce Type: cross Abstract: LLM agents increasingly draft messages on behalf of users, yet users routinely overshare sensitive information and disagree on what counts as private.

Shortcut Learning in Glomerular AI: Adversarial Penalties Hurt, Entropy Helps

SafetyDGX agent

arXiv:2604.07936v1 Announce Type: new Abstract: Stain variability is a pervasive source of distribution shift and potential shortcut learning in renal pathology AI. We ask whether lupus nephritis glom

Synthetic Homes: A Multimodal Generative AI Pipeline for Residential Building Data Generation under Data Scarcity

Model ReleasesDGX agent

arXiv:2509.09794v4 Announce Type: replace Abstract: Computational models have emerged as powerful tools for multi-scale energy modeling research at the building and urban scale, supporting data-driven

Transforming the Voice of the Customer: Large Language Models for Identifying Customer Needs

TutorialsDGX agent

arXiv:2503.01870v2 Announce Type: replace Abstract: Identifying customer needs (CNs) is fundamental to product innovation and marketing strategy. Yet for over thirty years, Voice-of-the-Customer (VOC)

9 Apr 2026

Apiiro launches command-line interface to bring AI-native security into software development workflows

Model ReleasesDGX agent

Application security posture management company Apiiro Ltd. today announced the launch of a new command-line interface designed to bring application security directly into artificial intelligence-driv

Is a backlash brewing? Rapid innovation in AI coding and agents may force push for enterprise order and control

Model ReleasesDGX agent

Artificial intelligence is proving to be a big bet for many companies across the enterprise landscape, and the gamblers are getting nervous. A survey of 2,400 global employees and C-suite leaders rele

Our company mission today is to give AI agents the highest-quality document context. The native open-source libs that agents have access to …

AgentsDGX agent

Our company mission today is to give AI agents the highest-quality document context. The native open-source libs that agents have access to (e.g. PyPDF) do naive text extraction. But this is incomplet

7 Apr 2026

This uses LiteParse, which is great for fast text search: https://developers.llamaindex.ai/liteparse/?utm_medium=social&utm_source=xjl If yo…

AgentsDGX agent

This uses LiteParse, which is great for fast text search: https://developers.llamaindex.ai/liteparse/?utm_medium=social&utm_source=xjl If you're interested in deeper VLM-enabled search, check out Llam

13 Aug 2026

Co-constructing sociotechnical AI governance: participatory system mapping using algorithm registers

SafetyDGX agent

arXiv:2608.12166v1 Announce Type: cross Abstract: Algorithm registers have been championed as a means of providing transparency on the use of algorithms in public services. Yet potential publics diffe

D3D-GEN: Robot-Aware Domain-Grounded Interactive 3D World Generation for Social Robotics

AgentsDGX agent

arXiv:2608.11876v1 Announce Type: new Abstract: Training and validation of Embodied AI for social navigation critically depends on realistic simulation environments, yet many current approaches fail t

Distribird: Literature-Informed Prior Distribution Design for Bayesian Model Calibration

Model ReleasesDGX agent

arXiv:2608.11210v1 Announce Type: new Abstract: Bayesian calibration of process-based models requires a prior distribution for each model parameter. Despite decades of methodological work, researchers

Evaluating LLM Generated Detection Rules in Cybersecurity

Model ReleasesDGX agent

arXiv:2509.16749v1 Announce Type: cross Abstract: LLMs are increasingly pervasive in the security environment, with limited measures of their effectiveness, which limits trust and usefulness to securi

Fixed Jinja chat template for Qwen 3.5, 3.6, and the new 3.8 release

Model ReleasesDGX agent

Qwen just released their first 3.8 model. The main addition in 3.8 is prompt-steered reasoning effort. You can tell the model how deeply to think by setting reasoning_effort to xhigh, medium, or low.

InfraBench: Evaluating Infrastructure Agents Across Layers, Lifecycle, and Risk

Model ReleasesDGX agent

arXiv:2608.11234v1 Announce Type: new Abstract: Managing modern computing infrastructure has become a steadily harder problem due to the ever-increasing complexity. Recent advances in AI agents create

Learning Loco-Manipulation From SMPC Demonstrations With Sparse Offline-to-Online RL

SafetyDGX agent

arXiv:2608.12063v1 Announce Type: cross Abstract: Integrating locomotion and manipulation is essential for robot autonomy, but scaling standard Reinforcement Learning (RL) to complex tasks is severely

Mechanist: AI as a Scientific Instrument for Discovering the Mechanisms of Intelligence

Model ReleasesDGX agent

arXiv:2608.12036v1 Announce Type: new Abstract: AI models have achieved remarkable success across diverse domains, yet the mechanisms underlying their capabilities and the risks they may pose remain p

NetlistBench: Evaluating LLM Reliability in SPICE Netlist Recognition and Manipulation

Model ReleasesDGX agent

arXiv:2608.12197v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used in circuit design workflows, yet their reliability on simulator-facing SPICE netlist recognition an

Principal Trait Analysis: Towards Deriving 'Skills' in Human-AI Collaboration

AgentsDGX agent

arXiv:2608.11460v1 Announce Type: new Abstract: Large Language Model-powered agents are increasingly used in the workplace via human-artificial intelligence (AI) collaboration. In this new era of work

Self-evolving network verifiers

AgentsDGX agent

arXiv:2608.11340v1 Announce Type: cross Abstract: Symbolic network verifiers can reason about correctness across vast spaces of routing inputs and failures, but only for the protocols and features an

StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2608.11671v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models can follow instructions and manipulate objects, but their performance often collapses out of distribution (OOD), whe

Towards Query-Agnostic RAG Evaluation via Query Coverage and Claim Verifiability

Model ReleasesDGX agent

arXiv:2608.11238v1 Announce Type: new Abstract: Retrieval-augmented generation improves the factuality of large language models by grounding responses in retrieved evidence, yet existing evaluation fr

12 Aug 2026

A Gateway Architecture for Enterprise MCP Authentication: Unifying Heterogeneous Auth, Identity Delegation, and the User / Non-User Persona Problem

AgentsDGX agent

arXiv:2608.10760v1 Announce Type: cross Abstract: The Model Context Protocol (MCP) has become the de-facto interface for connecting LLM agents to enterprise tools, and adoption has been explosive: wit

Agentic Instruction Data Selection: Let DataMaster Interpret Your Intent

AgentsDGX agent

arXiv:2608.10579v1 Announce Type: new Abstract: Although existing instruction data selection methods have introduced various metrics, the inherent complexity of real-world datasets makes it impractica

APCReg: Anatomical-Prior-Guided Coarse-to-Fine CBCT--IOS Registration via Multi-View Projection and Reliability-Controlled Residual Correction

SafetyDGX agent

arXiv:2608.09993v1 Announce Type: cross Abstract: Registration between cone-beam computed tomography (CBCT) and intraoral scans (IOS) is essential for patient-specific surgical planning. However, disp

CHORUS: Complementary Experts for High-Coverage Testbench Stimulus Generation

Model ReleasesDGX agent

arXiv:2608.10090v1 Announce Type: new Abstract: Large language models (LLMs) have advanced code generation, where executable feedback provides a more reliable learning signal than textual imitation al

CodeRabbit bags $143M to help companies get a grip on the explosion of AI-generated code

AgentsDGX agent

CodeRabbit Inc., the creator of a popular tool that automatically reviews artificial intelligence-generated code, is becoming more ambitious after closing on its latest 143 million Series C round of f

Curate Before You Connect: Identity and Ontology Tagging in a Production Knowledge Graph

Model ReleasesDGX agent

arXiv:2608.10644v1 Announce Type: new Abstract: Extraction produces candidate entities and relationships; writing them into a graph is where identity is decided, and identity decisions are destructive

Edge Phoneme Recognition for Children's Speech through Age-Aware Training

Model ReleasesDGX agent

arXiv:2608.10206v1 Announce Type: new Abstract: Detecting phonemes from children's speech has historically been difficult due to the scarcity of training data, and unique characteristics of children's

Evaluation-Conditioned Training: Teaching Models to Generalize to Stronger Oversight Regimes

SafetyDGX agent

arXiv:2608.10209v1 Announce Type: new Abstract: Feedback signals used to train Large Language Models (LLMs) are the primary driver of their behavior and our main lever for instilling alignment with hu

Google unveils the $399 Pixel Watch 5 with a satin pyrite case finish, offline Gemini, proactive AI suggestions, better GPS maps, and insulin resistance trends (Victoria Song/The Verge)

Model ReleasesDGX agent

Victoria Song / The Verge: Google unveils the 399 Pixel Watch 5 with a satin pyrite case finish, offline Gemini, proactive AI suggestions, better GPS maps, and insulin resistance trends — The 399 Goog

How to Dogfood Your AI Chat Agent: A Three-Layer Evaluation Framework with Goal-Directed NPC Simulation

AgentsDGX agent

arXiv:2608.09939v1 Announce Type: cross Abstract: Production teams deploying LLM chat agents face a specific quality assurance gap: existing evaluation tools test individual responses or simulate soci

HSSBench: Benchmarking Humanities and Social Sciences Ability for Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2506.03922v4 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) have demonstrated significant potential to advance a broad range of domains. However, current benchma

***Lights, inference, action…*** I’m so happy to share that @sequoia has led the Seed in @previewio. Developers are flying in magical AI-nat…

ApplicationsDGX agent

***Lights, inference, action…*** I’m so happy to share that @sequoia has led the Seed in @previewio. Developers are flying in magical AI-native editors. But creative tooling is still stuck in the pre-

No Free Labels: Limitations of LLM-as-a-Judge Without Human Grounding

Model ReleasesDGX agent

arXiv:2503.05061v3 Announce Type: replace Abstract: Reliable evaluation of large language models (LLMs) is critical as their deployment rapidly expands, particularly in high-stakes domains such as bus

Order Matters: LVLMs as Judges for Temporal Reasoning in Image Sequences

ResearchDGX agent

arXiv:2608.10908v1 Announce Type: cross Abstract: As generative multimedia evolves from static image synthesis to complex, interleaved visual narratives, a foundational bottleneck has emerged: the jud

PolypVision: A Three-Stage Hierarchical Deep Learning Framework for Classification and Segmentation of Colorectal Polyps

ResearchDGX agent

arXiv:2608.10649v1 Announce Type: new Abstract: Colorectal cancer (CRC) remains one of the leading causes of cancer-related mortality worldwide, predominantly arising from precancerous polyps. Accurat

Recovering Wasted Compute in Autoresearch Agents

AgentsDGX agent

arXiv:2608.10424v1 Announce Type: new Abstract: A slew of recent works develop agents for solving research problems end-to-end, a paradigm increasingly referred to as autoresearch. Such agents have in

ReLTEx: Reliable LLM-based Taxonomy Expansion

Model ReleasesDGX agent

arXiv:2608.10970v1 Announce Type: cross Abstract: Recent advances in Large Language Models (LLMs) have demonstrated strong capabilities in generating semantically relevant concepts and relations, maki

Significance and Stability Analysis of Gene-Environment Interaction using GxEStat

Model ReleasesDGX agent

arXiv:2604.03337v2 Announce Type: replace Abstract: Genotype-environment (GxE) interactions can influence the performance of genotypes across diverse environments, limiting the reliability of genotype

SinD 2.0: A Multi-City UAV Dataset with Semantic Risk Annotations for SOTIF-Oriented Safety Validation at Signalized Intersections

Model ReleasesDGX agent

arXiv:2607.16943v2 Announce Type: replace Abstract: Safety validation at signalized intersections remains a critical bottleneck for the deployment of autonomous driving systems (ADS), as these scenari

TACTICL: Task-Aware Compression of Tabular ICL Models

Model ReleasesDGX agent

arXiv:2608.10837v1 Announce Type: cross Abstract: The strong performance of foundation models for tabular tasks comes at substantial inference costs. Distilling models into task-specific architectures

The CASE Framework: A Multi-Disciplinary Control Architecture for Governing Enterprise Agentic AI

SafetyDGX agent

arXiv:2608.10153v1 Announce Type: new Abstract: Enterprises are deploying autonomous AI agents faster than they can govern them, and prevailing approaches stretch a single discipline, typically DevSec

Together Serverless Inference gives developers a managed, high-throughput path for running Qwen3.8-2.4T-A95B across coding and agentic workl…

AgentsDGX agent

Together Serverless Inference gives developers a managed, high-throughput path for running Qwen3.8-2.4T-A95B across coding and agentic workloads. Start building: https://www.together.ai/models/qwen3-8

VidForensics-M1: Meta-Detection Reinforcement Learning with Verifiable Temporal Grounding for AI-Generated Video Forensics

Local AiDGX agent

arXiv:2608.11201v1 Announce Type: new Abstract: Recent advances in video generation models have significantly improved the realism of synthetic videos, blurring the boundary between generated and auth

11 Aug 2026

A Grounded and Decomposed Framework for Relation-Level Hallucination Evaluation in Abstractive Summarization

AgentsDGX agent

arXiv:2608.08180v1 Announce Type: cross Abstract: Abstractive text summarization systems frequently generate fluent yet unfaithful summaries by fabricating or distorting relationships between entities

A Mixed-Stiffness Anthropomimetic Fingertip Broadens the Operating Range for Coin Grasping

ResearchDGX agent

arXiv:2608.07887v1 Announce Type: new Abstract: Robotic grasping of thin, flat objects such as coins on hard surfaces remains challenging because conventional methods require reorienting the object, a

A Unified Framework for Dynamic Reward Shaping in Reinforcement Learning

SafetyDGX agent

arXiv:2608.08158v1 Announce Type: new Abstract: Sparse, delayed, and weakly informative rewards remain central obstacles to efficient reinforcement learning. Reward shaping addresses these limitations

ACEvo: Adversarial Co-Evolution of Problem Distributions and Solvers for Combinatorial Optimization

Model ReleasesDGX agent

arXiv:2506.02594v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly used to synthesize heuristic programs, yet most existing pipelines optimize solvers against fixed benc

AIVV: Neuro-Symbolic LLM Agent-Integrated Verification and Validation for Trustworthy Autonomous Systems

AgentsDGX agent

arXiv:2604.02478v2 Announce Type: replace Abstract: Deep learning models excel at detecting anomaly patterns in normal data. However, they do not provide a direct solution for anomaly classification a

Beyond cognacy

SafetyDGX agent

arXiv:2507.03005v3 Announce Type: replace Abstract: Computational phylogenetics has become an established tool in historical linguistics, with many language families now analyzed using likelihood-base

BibTeX Citation Errors in Scientific Publishing Agents: Evaluation and Mitigation

Model ReleasesDGX agent

arXiv:2604.03159v2 Announce Type: replace-cross Abstract: Large language models with web search are increasingly used in scientific publishing agents, yet they produce BibTeX entries with pervasive fi

Coupled Graph--Policy Distillation for Personalized Medication Safety in Older Adults with Multimorbidity

Model ReleasesDGX agent

arXiv:2608.09443v1 Announce Type: new Abstract: Large language model (LLM) agents can support medication review between clinical visits, but safe choices for older adults with multimorbidity depend on

← Previous
1…4445464748…84
Next →