AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
10,012 results
Model Releases

A Fresh Look at Lamarckian Evolution and the Baldwin Effect

DGX agent

arXiv:2605.28703v1 Announce Type: cross Abstract: Baldwinian and Lamarckian evolution have existed for a long time in evolutionary algorithms (EAs) without ever dominating the academic literature or p

model-releasesarxiv-cs-ai
28 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

A Methodology to Assess Power Modeling in Energy-Aware Federated Learning on Heterogeneous Mobile Devices

DGX agent

arXiv:2605.27601v1 Announce Type: cross Abstract: Estimating CPU power on heterogeneous ARM-based commodity devices is challenging due to limited access to CPU's voltage domains. As a result, state-of

researcharxiv-cs-lg
28 May 2026
Research

A New Era of Innovation: Google Research at I/O 2026

DGX agent

Google Research at I/O 2026 showcased new Gemini AI models including Gemini Omni, which can create content from any input starting with video, and Gemini 3.5 Flash, combining frontier intelligence wit

researchgoogle-research
28 May 2026
Safety

A Policy-Driven Runtime Layer for Agentic LLM Serving

DGX agent

arXiv:2605.27744v1 Announce Type: new Abstract: Multi-agent LLM systems have become the dominant production workload, but the serving stack was not built for them. The agent framework above knows agen

safetyarxiv-cs-ai
28 May 2026
Research

A Systematic Evaluation of Retrieval-Augmented Generation and Language Models for Space Operations

DGX agent

arXiv:2605.27444v1 Announce Type: cross Abstract: The rapid expansion of space activities has led to an unprecedented accumulation of technical documentation, operational guidelines, and scientific li

researcharxiv-cs-ai
28 May 2026
Model Releases

A Unified Framework for the Evaluation of LLM Agentic Capabilities

DGX agent

arXiv:2605.27898v1 Announce Type: new Abstract: As LLMs are increasingly deployed as agents, reliable assessment of their agentic capabilities has become essential. However, reported benchmark scores

model-releasesarxiv-cs-ai
28 May 2026
Safety

Affective Music Recommendation: A Rollout-Based World Model for Offline Preference Optimization

DGX agent

arXiv:2605.28810v1 Announce Type: new Abstract: Functional music applications, from consumer focus and sleep aids to clinical interventions, share a distinctive recommendation problem: success is defi

safetyarxiv-cs-lg
28 May 2026
Agents

An Operator-Based Approach to STL

DGX agent

arXiv:2605.28092v1 Announce Type: new Abstract: Signal Temporal Logic (STL), has recently seen extensive development, owing to its rich expressivenes for autonomous planning and control. Nevertheless,

agentsarxiv-cs-ro
28 May 2026
Safety

Automated Estimation of Impact Time, Impact Location, and Shuttlecock Speed in Badminton Smashes Using Event Cameras

DGX agent

arXiv:2605.28011v1 Announce Type: new Abstract: Quantifying impact phenomena in badminton smashes is important for evaluating both athletic performance and equipment; however, conventional measurement

safetyarxiv-cs-cv
28 May 2026
Model Releases

Benchmarks are Not Enough: RAMP for Runtime Assessing of Agentic Models in Production Systems

DGX agent

arXiv:2605.27492v1 Announce Type: cross Abstract: LLM agents are rapidly evolving from coding assistants into autonomous software engineering systems. However, existing evaluation methodologies remain

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Beyond being fast, LiteParse is designed to provide highly accurate, semantically coherent text for LLM use. We benchmarked every open-sourc…

DGX agent

Beyond being fast, LiteParse is designed to provide highly accurate, semantically coherent text for LLM use. We benchmarked every open-source, model-free PDF parser on LLM QA tasks - from PyPDF to PyM

model-releasesjerry-liu--x
28 May 2026
Model Releases

Beyond Model Ranking: Predictability-Aligned Evaluation for Time Series Forecasting

DGX agent

arXiv:2509.23074v3 Announce Type: replace-cross Abstract: In the era of increasingly complex AI models for time series forecasting, progress is often measured by marginal improvements on benchmark lea

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Bias Leaves a Gradient Trail: Label-Free Bias Identification via Gradient Probes on Concept Decompositions

DGX agent

arXiv:2605.28780v1 Announce Type: new Abstract: Vision classifiers can exploit spurious correlations, achieving high in-distribution accuracy yet failing under distribution shift. Existing approaches

model-releasesarxiv-cs-cv
28 May 2026
Safety

Breaking the Script Barrier: Enabling Automatic Alignment for PoS-based ASR Error Analysis in Non-Latin Scripts

DGX agent

arXiv:2605.28438v1 Announce Type: new Abstract: Automatic Speech Recognition (ASR) systems are commonly evaluated using aggregate metrics such as Word Error Rate (WER), which do not capture the lingui

safetyarxiv-cs-cl
28 May 2026
Model Releases

Chrome Enterprise rolls out AI agents and automation to streamline security management

DGX agent

Google LLC today launched new enhancements to Chrome Enterprise, the company’s enterprise version of its Chrome browser, designed to provide greater administrative and security control to information

model-releasessiliconangle
28 May 2026
Model Releases

Claude Opus 4.8: 'a modest but tangible improvement'

DGX agent

Anthropic shipped Claude Opus 4.8 today. My favourite thing about it is this note in the release announcement: Users will find Opus 4.8 to be a modest but tangible improvement on its predecessor. Ther

model-releasessimon-willison
28 May 2026
Model Releases

Claude Opus 4.8 is now available in Windsurf and Devin CLI

DGX agent

Claude Opus 4.8 has been made available for use in both Windsurf (Cognition AI's code editor) and Devin CLI (their command-line interface). This release expands access to Anthropic's latest Claude mod

model-releasescognition-ai--x
28 May 2026
Safety

CPPO: Contrastive Perception Policy Optimization for VLM Agents

DGX agent

arXiv:2601.00501v2 Announce Type: replace Abstract: We introduce CPPO, a Contrastive Perception Policy Optimization method for finetuning vision--language models (VLMs). Reliable perception is a core

safetyarxiv-cs-cv
28 May 2026
Agents

Deploying a Hermes Agent with Fly, Modal, OpenRouter, & Cloudflare 02:43 Managed vs VPS 06:08 Architecture 15:14 Setup 22:23 Deployment 28:0…

DGX agent

Deploying a Hermes Agent with Fly, Modal, OpenRouter, & Cloudflare 02:43 Managed vs VPS 06:08 Architecture 15:14 Setup 22:23 Deployment 28:00 Access / OIDC 39:57 Hermes and Open WebUI 52:07 Cloudflare

agentsnous-research--x
28 May 2026
Agents

Do Agents Need Semantic Metadata? A Comparative Study in Agentic Data Retrieval

DGX agent

arXiv:2605.28787v1 Announce Type: cross Abstract: In the era of autonomous agents, machine-actionable data is critical for data-driven workflows. For more than a decade, semantic metadata like schema.

agentsarxiv-cs-ai
28 May 2026
Model Releases

Do Agents Think Deeper? A Mechanistic Investigation of Layer-Wise Dynamics in Sequential Planning

DGX agent

arXiv:2605.27935v1 Announce Type: new Abstract: Recent mechanistic studies suggest that large language models (LLMs) may utilize their depth inefficiently in standard single-turn tasks. Whether this s

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

DynaSchedBench: Calibrated Dynamic Scheduling Benchmarks and Observability Paradox in LLM-based Scheduling Agents

DGX agent

arXiv:2605.27566v1 Announce Type: new Abstract: Progress in neural combinatorial optimization for Dynamic Flexible Job Shop Scheduling Problem (DFJSP) is currently hindered by a methodological tension

model-releasesarxiv-cs-ai
28 May 2026
Agents

E^3-Agent: An Executable and Evolving Agent for Resource Management of Edge Generative Inference

DGX agent

arXiv:2605.27428v1 Announce Type: new Abstract: Edge deployments of generative inference increasingly face two practical realities: per-device per-model performance is often unknown at deployment time

agentsarxiv-cs-lg
28 May 2026
Research

EchoAvatar: Real-time Generative Avatar Animation from Audio Streams

DGX agent

arXiv:2605.28272v1 Announce Type: new Abstract: Real-time synthesis of high-fidelity 3D character motion from audio is a pivotal component for next-generation interactive avatars and virtual assistant

researcharxiv-cs-cv
28 May 2026
Research

Eliot: Interactively nderline{E}xploring Fast-Changing Scientific nderline{Li}terature Trends with nderline{O}nline Danderline{t}a and Learning

DGX agent

arXiv:2605.27610v1 Announce Type: cross Abstract: The rapid growth of scientific publishing has made it increasingly difficult to track how fast-moving areas evolve. Search engines and LLM-based assis

researcharxiv-cs-ai
28 May 2026
Agents

Evals shape agent behavior. Every eval is a vector that shifts the behavior of your agentic system. More evals ≠ better agents. Instead, bui…

DGX agent

Evals shape agent behavior. Every eval is a vector that shifts the behavior of your agentic system. More evals ≠ better agents. Instead, build targeted evals that reflect desired behaviors in producti

agentsharrison-chase--x
28 May 2026
Model Releases

Evolving Dataflow to process massive datasets for machine learning

DGX agent

Google created MapReduce more than 20 years ago to solve the scaling problems in data processing that the then young company was running into. The AI era that we are in now demands efficient, large-sc

model-releasesgoogle-cloud-ai
28 May 2026
Agents

Extrapolative Weight Averaging Reveals Correctness-Efficiency Frontiers in Code RL

DGX agent

arXiv:2605.28751v1 Announce Type: cross Abstract: Linear interpolation between fine-tuned checkpoints has been shown to trace the Pareto front between competing objectives, but whether extrapolative w

agentsarxiv-cs-ai
28 May 2026
Industry

FBI says Google engineer used internal search data to win $1.2M on Polymarket

DGX agent

Federal prosecutors charged a Google software engineer with making roughly $1.2 million in profits from bets on the prediction market platform Polymarket by using confidential insider information he l

industryars-technica
28 May 2026
Model Releases

ForestHG-Trace: Traceable Long-Horizon Ecological Reasoning over Large-Scale Forest Scenes

DGX agent

arXiv:2605.27590v1 Announce Type: new Abstract: Remote sensing question answering (RS-QA) often requires more than direct semantic prediction, especially in large-scale forest scenes where ecological

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

From Knowing to Doing: A Memory-Controlled Benchmark for LLM Trading Agents on Stock Markets

DGX agent

arXiv:2605.28359v1 Announce Type: new Abstract: Evaluating whether large language model (LLM) agents can profit in capital markets is increasingly framed as end-to-end trading: place an agent in a his

model-releasesarxiv-cs-ai
28 May 2026
Industry

Good value for money

DGX agent

Good value for money We gave Grok Build 0.1 one prompt: build a webhook delivery service in TypeScript, Bun, and SQLite. It planned it, built it, and shipped a working demo. Total cost: $1.65. Zero to

industryelon-musk--x
28 May 2026
Agents

How Endava builds an agentic organization with Codex

DGX agent

Endava, a software services company, leverages OpenAI's Codex to transform its organizational operations into an agentic model where AI agents autonomously handle tasks and decision-making. The approa

agentsopenai
28 May 2026
Applications

I have spoken to lots of companies that blew through their token budget for the year in months, but a half a billion dollars on internal emp…

DGX agent

Ethan Mollick discusses companies that rapidly depleted their annual token budgets within months, with some spending hundreds of millions of dollars on internal employee use of large language models.

applicationsethan-mollick--x
28 May 2026
Model Releases

I'm proud to share that @Glean has surpassed 300M ARR, just five months after crossing 200M and growing ~3x over the past 15 months. This …

DGX agent

I'm proud to share that @Glean has surpassed 300M ARR, just five months after crossing 200M and growing ~3x over the past 15 months. This is an exciting milestone for Glean, and it's a signal about wh

model-releasessonya-huang--x
28 May 2026
Safety

Insurance Pricing Optimization via Off-Policy Evaluation

DGX agent

arXiv:2605.28327v1 Announce Type: cross Abstract: Traditional insurance pricing relies on risk-based principles that ensure actuarial fairness and solvency but do not explicitly account for policyhold

safetyarxiv-cs-lg
28 May 2026
Model Releases

Interpretability-Guided Layer Selection over Subspace Projection: SAEs as Stethoscopes, Not Scalpels, for Raw Task Vector Model Editing

DGX agent

arXiv:2605.28649v1 Announce Type: cross Abstract: LLMs increasingly require surgical model editing to enhance domain-specific capabilities without incurring the computational cost or catastrophic forg

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

KSAFE-MM: A Multimodal Safety Benchmark via Localized Contextualization for Korean Cultural Risks

DGX agent

arXiv:2605.28013v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) exacerbate safety risks by introducing vulnerabilities across multiple modalities, such as language and vision.

model-releasesarxiv-cs-cl
28 May 2026
Agents

Large Language Models as Automatic Annotators and Annotation Adjudicators for Fine-Grained Opinion Analysis

DGX agent

arXiv:2601.16800v3 Announce Type: replace Abstract: Fine-grained opinion analysis of text provides a detailed understanding of expressed sentiments, including the addressed entity. Although this level

agentsarxiv-cs-cl
28 May 2026
Research

Learning Logical Operations for Arbitrary Quantum Error Correction Codes

DGX agent

arXiv:2605.28162v1 Announce Type: cross Abstract: Logical operations are essential for quantum computation within quantum error-correcting codes. However, discovering their physical realizations is ch

researcharxiv-cs-lg
28 May 2026
Safety

Learning with Importance Weighted Variational Inference

DGX agent

arXiv:2410.12035v2 Announce Type: replace-cross Abstract: Several variational bounds involving importance weighting ideas generalize the Evidence Lower BOund (ELBO) for marginal likelihood optimizatio

safetyarxiv-cs-lg
28 May 2026
Model Releases

Local Reachy Mini conversations wireless looks like magic! You can bring your friend around the house and get WOW effect from anyone! 🔥 Tha…

DGX agent

Local Reachy Mini conversations wireless looks like magic! You can bring your friend around the house and get WOW effect from anyone! 🔥 Thanks @andimarafioti for the blog post on how to set this up: h

model-releasesclem-delangue--x
28 May 2026
Safety

Mathematical Modelling of Ethical AI Use in Higher Education: A Coordination Game Framework for Future-Facing Learning

DGX agent

arXiv:2605.27400v1 Announce Type: cross Abstract: The rapid uptake of generative artificial intelligence (AI) in higher education is reshaping assessment practices and intensifying concerns around aca

safetyarxiv-cs-ai
28 May 2026
Applications

models being conscious would be harmful for humanity. it would encroach on our status and dignity. it would limit the type of things we can …

DGX agent

models being conscious would be harmful for humanity. it would encroach on our status and dignity. it would limit the type of things we can do with them and use them for. it would vastly accelerate hu

applicationsethan-mollick--x
28 May 2026
Agents

Multi-Agent LLM-based Metamorphic Testing for REST APIs

DGX agent

arXiv:2605.28321v1 Announce Type: cross Abstract: As REST APIs become an increasingly significant part of software systems, their validation is becoming more critical. Hence, testing and uncovering un

agentsarxiv-cs-ai
28 May 2026
Model Releases

On the Subgaussianity of Quantized Linear Maps: An AI-Assisted Note

DGX agent

arXiv:2605.27563v1 Announce Type: cross Abstract: This short note presents a dimension-independent subgaussian concentration bound for Gaussian vectors under coordinate-wise nonlinear mappings. Discov

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Pittsburgh-based Gray Swan, which stress-tests AI models for top frontier AI labs, raised a 40M Series A at a 200M valuation co-led by Wing VC and Madrona (Rashi Shrivastava/Forbes)

DGX agent

Rashi Shrivastava / Forbes: Pittsburgh-based Gray Swan, which stress-tests AI models for top frontier AI labs, raised a 40M Series A at a 200M valuation co-led by Wing VC and Madrona — Gray Swan works

model-releasestechmeme
28 May 2026
Model Releases

Plant, Persist, Trigger: Sleeper Attack on Large Language Model Agents

DGX agent

arXiv:2605.28201v1 Announce Type: new Abstract: Large Language Model (LLM) agents remain vulnerable to safety threats from the external environment, where attackers inject adversarial content into ext

model-releasesarxiv-cs-ai
28 May 2026
← Previous
1…166167168169170…209
Next →