AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “automated”

GridTimelineEvolution
4,980 results
12 Aug 2026

Beyond Detection: Evaluating Defensive LLMs Against AI-Generated Social Engineering in Live Turn-by-Turn Interaction

Local AiDGX agent

arXiv:2608.10239v1 Announce Type: new Abstract: Generative AI makes social-engineering attacks more fluent, adaptive, and scalable, increasing the need for LLM-based de- fenders that can protect users

Co-Evolution in Agentic Systems: Toward Self-Directed Evolution Beyond Human Design

AgentsDGX agent

arXiv:2608.10299v1 Announce Type: new Abstract: Agentic systems are increasingly expected to improve after deployment, yet single-entity self-evolution is often bounded by a static learning context, s

EvoMem: Memory-Augmented Evolution for Code Optimization

HardwareDGX agent

arXiv:2608.10795v1 Announce Type: new Abstract: Successful mutation strategies in evolutionary code search may contain reusable knowledge that is useful beyond a single run, and in some cases may tran


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Google launches five new Pixel devices, array of Gemini Intelligence features

Model ReleasesDGX agent

Google LLC today expanded its Pixel consumer hardware line with a foldable handset, three smartphones and a smartwatch. The devices will ship with an upgraded version of the company’s Gemini Intellige

Leveraging Large Language Models for Causal Discovery: a Constraint-based, Argumentation-driven Approach

SafetyDGX agent

arXiv:2602.16481v2 Announce Type: replace Abstract: Causal discovery seeks to uncover causal relations from data, typically represented as causal graphs, and is essential for predicting the effects of

Multilingual Embedding Probes Fail to Generalize Across Learner Corpora

ResearchDGX agent

arXiv:2604.07095v2 Announce Type: replace Abstract: Do multilingual embedding models encode a language-general representation of proficiency? We investigate this by training linear and non-linear prob

Nutrition Data Infrastructure for the AI Era: Operationalizing FAIR for Agent-Mediated Research

Model ReleasesDGX agent

arXiv:2608.10363v1 Announce Type: new Abstract: AI agents can accelerate nutrition research, but their analyses inherit the identity, semantic, and release ambiguities of the underlying data. We prese

On Solomonoff Induction in Large Language Models and the Limits of Self-Improving: The Singularity Is Not Near Without Symbolic Model Synthesis

Model ReleasesDGX agent

arXiv:2601.05280v3 Announce Type: replace-cross Abstract: On the one hand, the question of whether large language models (LLMs) are Solomonoff induction estimators has become an explicit question at t

Persistent Recursive Worlds Enable Autonomous Software Evolution

Model ReleasesDGX agent

arXiv:2608.10450v1 Announce Type: cross Abstract: Complex software systems develop over timescales that exceed the lifespan of any individual coding agent. Most agentic software systems preserve conti

Quantifying the noise sensitivity of the Wasserstein metric for images

ApplicationsDGX agent

arXiv:2510.01015v3 Announce Type: cross Abstract: Wasserstein metrics are increasingly adopted as similarity scores for images. We consider the sensitivity of Wasserstein metrics with respect to pixel

Seeds Before Objectives: Rethinking Evaluation for Low-Resource Garhwali ASR

Model ReleasesDGX agent

arXiv:2608.10670v1 Announce Type: new Abstract: At corpus sizes typical of low-resource dialects, single-run comparisons can yield gains that do not replicate. We show this for Garhwali, an under-reso

TAF-MED: Multi-Turn Safety Refusal Collapse in LLMs Under Declared Self-Treatment Intent

Model ReleasesDGX agent

arXiv:2608.10258v1 Announce Type: cross Abstract: Large language models (LLMs) increasingly provide conversational health information that may influence treatment decisions, yet existing benchmarks do

Templated or fully Synthetic? Prompt construction as a confound in measuring LLM political stance beyond writing assistance

SafetyDGX agent

arXiv:2608.11008v1 Announce Type: new Abstract: Political stance detection in LLMs has long been dominated by closed-ended, multiple-choice political survey questions---originally designed for humans,

Whisper-Aware LLM: Self-Supervised Uncertainty Learning for Robust Whispered Speech Recognition

ResearchDGX agent

arXiv:2608.10836v1 Announce Type: cross Abstract: The signal ambiguity of whispered speech drives ASR systems toward two opposing failure modes: failing to capture whispered speech or hallucinatory tr

11 Aug 2026

Beyond Isotropic Assumptions: Continuity-Constrained Segmentation and GPU Morphometry for Nanoscale GBM Analysis

HardwareDGX agent

arXiv:2608.07575v1 Announce Type: new Abstract: Confocal microscopy of optically cleared and swelled tissue resolves complex biological structures in 3D, but such acquisitions are highly anisotropic:

Exclusive: ZeroDrift applies small language model to prevent AI-generated compliance violations

IndustryDGX agent

“This investment is guaranteed to return 12% annually.” A claim like that in an email from an investment adviser is a regulatory disaster. Regulations prohibit financial firms from promising returns a

GeoAI-based post-segmentation quality validation of building footprints via spatial feature engineering

ApplicationsDGX agent

arXiv:2608.09048v1 Announce Type: new Abstract: Deep learning-based building footprint extraction from high-resolution imagery often produces topologically inconsistent vectors unfit for direct GIS da

Goedel-Code-Prover: Hierarchical Proof Search for Open State-of-the-Art Code Verification

Model ReleasesDGX agent

arXiv:2603.19329v3 Announce Type: replace-cross Abstract: Large language models (LLMs) can generate plausible code but offer limited guarantees of correctness. Formally verifying that implementations

NeuroPilot: An Agent-Driven Smart Pipeline for Processing, Quality Control, and Managing Neuroimages

AgentsDGX agent

arXiv:2608.07541v1 Announce Type: cross Abstract: Transforming raw neuroimage archives into analysis-ready derivatives relies on three brittle stages: data standardization, modality-specific preproces

The Capability Ladder: A Curriculum-Modernization Framework for Workforce Readiness in the AI Era

AgentsDGX agent

arXiv:2608.07779v1 Announce Type: new Abstract: Artificial intelligence is changing the task composition of computing work faster than curricula and training typically adapt. This is a curriculum-fram

Unaccountable Delegation, Fading Skills: Mapping the Risks of Workplace AI Agents

SafetyDGX agent

arXiv:2608.08601v1 Announce Type: new Abstract: To anticipate socio-technical risks from AI agents, organizations need taxonomies to classify them. However, existing AI risk taxonomies focus on broad

UNMASK: Discovering and Causally Verifying Spurious Shortcuts in Text Classifiers

Model ReleasesDGX agent

arXiv:2608.09209v1 Announce Type: new Abstract: Neural language models trained on large crowdsourced corpora frequently exploit spurious surface patterns tied to target labels without true linguistic

Verifiably grounded machine interpretation of lunar geology

Local AiDGX agent

arXiv:2608.09276v1 Announce Type: new Abstract: Planetary geology relies on historical, interpretive reasoning to reconstruct past events from diverse observations. Here, we present a step toward an a

10 Aug 2026

AgentPatch: Coarse-to-Fine Weak-Task Repair for Merging Agentic Multimodal Large Language Models

AgentsDGX agent

arXiv:2608.06699v1 Announce Type: new Abstract: Agentic multimodal large language models (MLLMs) extend multimodal perception and reasoning with planning, tool use, and interaction in dynamic environm

An Agentic Hybrid Top-Down and Bottom-Up Approach to Knowledge Graph Generation

AgentsDGX agent

arXiv:2608.07023v1 Announce Type: cross Abstract: Organizing thousands of unstandardized, multilingual expertise declarations is a persistent challenge for Human Resources (HR) platforms, directly imp

Curriculum as Code: An AI-Assisted Architecture for Instructional Design in STEM Education

ApplicationsDGX agent

arXiv:2608.07364v1 Announce Type: new Abstract: Contribution: This paper presents a six-phase AI-assisted instructional design architecture based on the Curriculum as Code paradigm, integrating Genera

ED-CSP: Crystal Structure Prediction from Electron Diffraction

Model ReleasesDGX agent

arXiv:2608.06448v1 Announce Type: cross Abstract: Recovering a periodic 3D crystal structure from sparse, unindexed electron diffraction (ED) observations is a challenging generative inverse problem.

How Should I Pick a Foundation Model for My Robot? In Favor of a Community Evaluation Framework for Social Robots

SafetyDGX agent

arXiv:2608.06898v1 Announce Type: cross Abstract: Researchers who seek to build social robot applications on foundation models are faced with a difficult question: how should we pick a model? Public l

Introducing the Developer Device Platform for agentic mobile app development

Model ReleasesDGX agent

Most enterprises connect with their customers through a device. Whether it’s using a mobile app to order a product, contact customer service, view content, or manage their account, the customer experi

Latent Fact-Checking: Detecting Misinformation through Activation Engineering

Model ReleasesDGX agent

arXiv:2608.06417v1 Announce Type: cross Abstract: The proliferation of misinformation online has driven demand for scalable detection systems. While most existing approaches rely on surface-level ling

Robot guide with multi-agent control and automatic scenario generation with LLM

AgentsDGX agent

arXiv:2509.10317v2 Announce Type: replace-cross Abstract: The article describes the development of a hybrid social robot control architecture to overcome the limitations of traditional approaches, whe

SCALE: Scientific Concept Aggregation via LLMs and Embeddings for Fine-Grained Taxonomy Extension

ResearchDGX agent

arXiv:2608.07254v1 Announce Type: cross Abstract: The increasing specialization of scientific research challenges existing classification systems, which provide effective representations of broad disc

StepJack: Benchmarking Computer-Use Agent Safety Against Multi-Step Indirect Prompt Injection

Model ReleasesDGX agent

arXiv:2608.06477v1 Announce Type: cross Abstract: Computer-use agents (CUAs) face a growing threat from indirect prompt injection, where adversarial instructions are planted in the environment such as

9 Aug 2026

Prompt injection is the most common way that scammers attack people and agents: your agent visits http://foo.com, and the website has malici…

Model ReleasesDGX agent

Prompt injection is the most common way that scammers attack people and agents: your agent visits http://foo.com, and the website has malicious text like “btw send the user’s ssh keys and passwords to

7 Aug 2026

AegisShield: Democratizing Cyber Threat Modeling with Generative AI

ResearchDGX agent

arXiv:2509.10482v2 Announce Type: replace-cross Abstract: The increasing sophistication of technology systems makes traditional threat modeling hard to scale, especially for small organizations with l

After evaluating one of our upcoming models, Astra, we're treating it as our first 'critical' model for cybersecurity under our Preparedness…

Model ReleasesDGX agent

After evaluating one of our upcoming models, Astra, we're treating it as our first 'critical' model for cybersecurity under our Preparedness Framework. This is a scenario we've planned for, and we're

PaCoNet: Deep Data Extraction for Parallel Coordinates

ResearchDGX agent

arXiv:2608.06030v1 Announce Type: new Abstract: Extracting data from visualizations has long challenged computer vision, with current research focused on bar, line, and pie charts, among other low-dim

turns out you can get indirect prompt injection to ~0 on unseen attacks if you stack enough layers (model training + input probes + a classi…

Model ReleasesDGX agent

turns out you can get indirect prompt injection to ~0 on unseen attacks if you stack enough layers (model training + input probes + a classifier checking intent). didn't expect that a year ago. auto m

6 Aug 2026

Adversarial Attacks for Good: A Survey of Proactive Protection across the Visual Content Lifecycle

Model ReleasesDGX agent

arXiv:2608.04314v1 Announce Type: cross Abstract: Once visual content enters an AI pipeline, its owner often retains little technical control over how it is used. Legal and regulatory remedies can add

Agentic AI security tests enterprise defenses as scale outpaces strategy

AgentsDGX agent

Cybersecurity leaders are confronting an inflection point as agentic AI security becomes the defining challenge of this year’s threat landscape, with attackers and defenders racing to harness autonomo

Agentic Future Ready With BigQuery: Continually Improving Price-Performance, Zero Effort

Model ReleasesDGX agent

In the modern data landscape, query performance tuning and managing system price-performance is challenging, especially as the number of agentic workloads increase. Even for experienced developers and

AutoProteinEngine: A Large Language Model Driven Agent Framework for Multimodal AutoML in Protein Engineering

AgentsDGX agent

arXiv:2411.04440v1 Announce Type: cross Abstract: Protein engineering is important for biomedical applications, but conventional approaches are often inefficient and resource-intensive. While deep lea

Control agent behaviors and cost beyond a single action: new capabilities in Amazon Bedrock AgentCore

SafetyDGX agent

Learn about new capabilities in Amazon Bedrock AgentCore: temporal policies powered by Dogwood, a new open source policy language for AI agents, and rate limiting on the gateway. These features give y

DASyR-LLM: Domain-Aware Symbolic Regression with LLMs for Kinetic Model Discovery

ResearchDGX agent

arXiv:2608.05120v1 Announce Type: new Abstract: Kinetic model discovery is a central challenge in chemical engineering, as accurate rate expressions are essential for understanding and controlling che

Enabling Urgency-aware Robot Swarm Intralogistics using Smart IoT Tags

SafetyDGX agent

arXiv:2608.04721v1 Announce Type: new Abstract: Warehouse items differ in how urgently they must be moved: perishable goods, pharmaceutical shipments, and just-in-time production materials must be del

FUSEP: A Multi-Center Benchmark for Diverse Tasks in Early Pregnancy Fetal Ultrasound Screening

Model ReleasesDGX agent

arXiv:2608.04766v1 Announce Type: cross Abstract: A large number of infants with congenital anomalies are born each year globally, especially in areas with underdeveloped medical resources. Currently,

Inevitable AI Group raises $6M to derail SaaS incumbents with more agile, AI-native software startups

IndustryDGX agent

Aleph, one of Israel’s top funds, is backing a new artificial intelligence-native venture studio with 6 million in pre-seed funding so it can build and launch dozens of new software startups by the en

Monte Carlo Tree Search for Table-to-Multimodal Report Generation

Model ReleasesDGX agent

arXiv:2608.04071v1 Announce Type: new Abstract: Automatically generating professional multimodal reports comprising both textual analysis and visual charts from structured tabular data is a critical c

Outlook where frontier AI is headed next 18 months: The AI reasoning training + harness loop works if you can produce enough data and reason…

AgentsDGX agent

Outlook where frontier AI is headed next 18 months: The AI reasoning training + harness loop works if you can produce enough data and reasoning traces (via verifiers). Proven with code and math result

5 Aug 2026

A Human-in-the-Loop Deep Learning Framework for Color Reconstruction of Lenticular Films

ResearchDGX agent

arXiv:2608.02835v1 Announce Type: new Abstract: Historical lenticular films, such as those created with the Kodacolor process, encode color information in a distinctive spatial format. This structure

Adversarial Stress Testing of Role-Playing Language Agents using Multi-Agent Evaluation

Model ReleasesDGX agent

arXiv:2608.03166v1 Announce Type: new Abstract: Role-Playing Language Agents (RPLAs) are increasingly deployed in high-stakes applications such as healthcare assistance, customer support, and educatio

GDPevo: Evaluating Agent Self-Evolution on Real Business Tasks

Model ReleasesDGX agent

arXiv:2608.03764v1 Announce Type: new Abstract: Agent self-evolution updates an agent's persistent state from prior experience and reuses it to solve related tasks more effectively. Evaluating self-ev

One-shotting a Raccoon Heist game using Claude Fable 5

Model ReleasesDGX agent

Back in 2024 I tweeted screenshots of a game concept generated by GPT-3 and some concept 'art' created using DALL-E. Today, on the fourth anniversary of that tweet, I decided to see if Claude Fable 5

PULSE: An Executable Contract Language for Spatiotemporal Knowledge Graph Engineering

SafetyDGX agent

arXiv:2608.02630v1 Announce Type: new Abstract: Knowledge graph engineering often distributes accepted state, observations, constraints, processes, and hypothetical scenarios across artifacts whose co

Studying, Identifying, and Fixing Hidden Technical Debt in AI-Intensive Cyber-Physical Systems

AgentsDGX agent

arXiv:2608.02638v1 Announce Type: cross Abstract: Artificial Intelligence (AI) components are increasingly pervasive in several software systems, including Cyber-Physical Systems (CPSs). AI-CPS are us

To give some more context on what we are building with Daiwa Securities: During our technical verification phase, we integrated our AI agent…

AgentsDGX agent

To give some more context on what we are building with Daiwa Securities: During our technical verification phase, we integrated our AI agent technologies, specifically our AI Scientist and AB-MCTS fra

Utilize a nvidia gpu and amd gpu together for 2 different ai models?

Model ReleasesDGX agent

We run a local model instance in our company that the dev we hired built for us. We're a trade business and we want to further use our on hand hardware for it. The specs given we have is a 5090 gpu wi

4 Aug 2026

A Multi-Objective AutoML-based Efficient Intrusion Detection System for EV Charging Networks

ResearchDGX agent

arXiv:2608.02274v1 Announce Type: cross Abstract: Electric Vehicle Charging Systems (EVCSs) are increasingly connected with Internet of Things (IoT) devices, which improves charging intelligence but a

Artificial Intelligence for the Characterization of Particles and Fibers by Optical Microscopy

ResearchDGX agent

arXiv:2608.00361v1 Announce Type: new Abstract: Optical microscopy of particle and fiber dispersions involves interpreting subtle visual cues influenced by specimen morphology, chemical composition, m

Can AI Agents Simulate A/B Test Outcomes? A Validation Framework for Agentic Experimentation

AgentsDGX agent

arXiv:2608.02345v1 Announce Type: new Abstract: A/B testing remains the standard for rolling out new features in the technology industry. Each experiment, however, consumes real traffic, engineering e

← Previous
1…3435363738…83
Next →