AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “engineering”

GridTimelineEvolution
5,375 results
15 May 2026

Concurrency without Model Changes: Future-based Asynchronous Function Calling for LLMs

AgentsDGX agent

arXiv:2605.15077v1 Announce Type: cross Abstract: Function calling, also known as tool use, is a core capability of modern LLM agents but is typically constrained by synchronous execution semantics. U

GraphFlow: An Architecture for Formally Verifiable Visual Workflows Enabling Reliable Agentic AI Automation

Local AiDGX agent

arXiv:2605.14968v1 Announce Type: new Abstract: GraphFlow is a visual workflow system designed to improve the reliability of agentic AI automation in multi-step, mission-critical processes. In these w

Graphs of Research: Citation Evolution Graphs as Supervision for Research Idea Generation

ResearchDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.14790v1 Announce Type: cross Abstract: Research idea generation is the innovation-driving step of automated scientific research. Recently, large language models (LLMs) have shown potential

Grounded Continuation: A Linear-Time Runtime Verifier for LLM Conversations

AgentsDGX agent

arXiv:2605.14175v1 Announce Type: new Abstract: In long conversations, an LLM can produce a next utterance that sounds plausible but rests on premises the conversation has already abandoned. Context-m

IFPV: An Integrated Multi-Agent Framework for Generative Operational Planning and High-Fidelity Plan Verification

AgentsDGX agent

arXiv:2605.14851v1 Announce Type: cross Abstract: Operational plan generation and verification are critical for modern complex and rapidly changing battlefield environments, yet traditional generation

Measuring Google AI Overviews: Activation, Source Quality, Claim Fidelity, and Publisher Impact

ResearchDGX agent

arXiv:2605.14021v1 Announce Type: cross Abstract: Google AI Overviews (AIOs) are arguably the most widely encountered deployment of generative AI, reaching over 2 billion users who may not realize the

MediaClaw: Multimodal Intelligent-Agent Platform Technical Report

AgentsDGX agent

arXiv:2605.14771v1 Announce Type: new Abstract: MediaClaw is a multimodal agent platform built on the OpenClaw ecosystem. Its core design follows a three-layer architecture of unified abstraction, plu

MoZoo:Unleashing Video Diffusion power in animal fur and muscle simulation

Model ReleasesDGX agent

arXiv:2605.13857v1 Announce Type: cross Abstract: The creation of cinematic-quality animal effects necessitates the precise modeling of muscle and fur dynamics, a process that remains both labor-inten

Native Parallel Reasoner: Reasoning in Parallelism via Self-Distilled Reinforcement Learning

SafetyDGX agent

arXiv:2512.07461v3 Announce Type: replace Abstract: We introduce Native Parallel Reasoner (NPR), a teacher-free framework that enables Large Language Models (LLMs) to self-evolve genuine parallel reas

PEML: Parameter-efficient Multi-Task Learning with Optimized Continuous Prompts

Model ReleasesDGX agent

arXiv:2605.14055v1 Announce Type: cross Abstract: Parameter-Efficient Fine-Tuning (PEFT) is widely used for adapting Large Language Models (LLMs) for various tasks. Recently, there has been an increas

PreFT: Prefill-only finetuning for efficient inference

Model ReleasesDGX agent

arXiv:2605.14217v1 Announce Type: cross Abstract: Large language models can now be personalised efficiently at scale using parameter efficient finetuning methods (PEFTs), but serving user-specific PEF

Randomized Atomic Feature Models for Physics-Informed Identification of Dynamic Systems

ResearchDGX agent

arXiv:2605.14351v1 Announce Type: cross Abstract: We present a physics-informed framework for system identification based on randomized stable atomic features. Impulse responses are represented as ran

SceneParser: Hierarchical Scene Parsing for Visual Semantics Understanding

Model ReleasesDGX agent

arXiv:2605.14923v1 Announce Type: new Abstract: General scene perception has progressed from object recognition toward open-vocabulary grounding, part localization, and affordance prediction. Yet thes

SCRWKV: Ultra-Compact Structure-Calibrated Vision-RWKV for Topological Crack Segmentation

ApplicationsDGX agent

arXiv:2605.14926v1 Announce Type: new Abstract: Achieving pixel-level accurate segmentation of structural cracks across diverse scenarios remains a formidable challenge. Existing methods face signific

Seed3D 2.0: Advancing High-Fidelity Simulation-Ready 3D Content Generation

Local AiDGX agent

arXiv:2605.13862v1 Announce Type: cross Abstract: We present Seed3D 2.0, an advanced 3D content generation system built on Seed3D 1.0, with substantial improvements across generation fidelity, simulat

SimPersona: Learning Discrete Buyer Personas from Raw Clickstreams for Grounded E-Commerce Agents

SafetyDGX agent

arXiv:2605.14205v1 Announce Type: new Abstract: LLM-based web agents can navigate live storefronts, yet they often collapse to a single 'average buyer' policy, failing to capture the heterogeneous and

The most revealing thing about this AI leadership paper is that it reads less like a vision for innovation and more like a glossy whitepaper…

Model ReleasesDGX agent

The most revealing thing about this AI leadership paper is that it reads less like a vision for innovation and more like a glossy whitepaper for a 21st century East India Company. Every generation of

Unlocking Complex Visual Generation via Closed-Loop Verified Reasoning

Model ReleasesDGX agent

arXiv:2605.14876v1 Announce Type: cross Abstract: Despite rapid advancements, current text-to-image (T2I) models predominantly rely on a single-step generation paradigm, which struggles with complex s

Watermarking Game-Playing Agents in Perfect-Information Extensive-Form Games

ResearchDGX agent

arXiv:2605.14283v1 Announce Type: cross Abstract: Watermarking techniques for large language models (LLMs), which encode hidden information in the output so its source can be verified, have gained sig

What Do EEG Foundation Models Capture from Human Brain Signals?

TutorialsDGX agent

arXiv:2605.11410v2 Announce Type: replace Abstract: Clinical electroencephalogram (EEG) analysis rests on a hand-crafted feature catalog refined over decades, e.g., band power, connectivity, complexit

With Devin, the migration finished 5.2x faster than projected. Full story: https://devin.ai/customers/angellist

AgentsDGX agent

Cognition AI reports that Devin, their AI coding assistant, enabled AngelList to complete a data migration 5.2 times faster than their original project timeline. The case study highlights Devin's effe

Wow! What a week! Lots of new things to learn - new ideas to form The complete Agent Development Life Cycle 🎯 ❤️ SF but can’t wait to get h…

AgentsDGX agent

Wow! What a week! Lots of new things to learn - new ideas to form The complete Agent Development Life Cycle 🎯 ❤️ SF but can’t wait to get home and back to building! Thanks for the warm welcome and fun

14 May 2026

A Multi-Agent Orchestration Framework for Venture Capital Due Diligence

AgentsDGX agent

arXiv:2605.13110v1 Announce Type: cross Abstract: We present a fully automated multi-agent framework for corporate due diligence and market analysis in venture capital. The system runs on an event-dri

after 15 years of waiting, the developers of singapore gave up on waiting for the government to get the tech sector going and finally brough…

ToolsDGX agent

after 15 years of waiting, the developers of singapore gave up on waiting for the government to get the tech sector going and finally brought SF to SG. great showings from @daytonaio @usetusk @arizeai

AgenticAITA: A Proof-Of-Concept About Deliberative Multi-Agent Reasoning for Autonomous Trading Systems

SafetyDGX agent

arXiv:2605.12532v1 Announce Type: cross Abstract: Conventional algorithmic trading systems are grounded in deterministic heuristics or offline-trained statistical models that cannot adapt to the seman

b9140

Local AiDGX agent

The search did not return specific information about the b9140 release. Based on the context, b9140 is a specific build number from the llama.cpp project releases. llama.cpp is an LLM inference implem

BEHAVE: A Hybrid AI Framework for Real-Time Modeling of Collective Human Dynamics

SafetyDGX agent

arXiv:2605.12730v1 Announce Type: new Abstract: Existing AI systems for modeling human behavior operate at the level of individuals or detect events after they occur. As a result, they systematically

British inference chip startup Fractile bags $220M to accelerate token consumption

HardwareDGX agent

U.K.-based artificial intelligence inference chip startup Fractile Ltd. said today it has closed on a 220 million Series B round of funding. The company was founded in 2022 by the Oxford University-tr

CHAL: Council of Hierarchical Agentic Language

AgentsDGX agent

arXiv:2605.12718v1 Announce Type: new Abstract: Multi-agent debate has emerged as a promising approach for improving LLM reasoning on ground-truth tasks, yet current methodologies face certain structu

Constraint-Aware Flow Matching: Decision Aligned End-to-End Training for Constrained Sampling

ApplicationsDGX agent

arXiv:2605.12754v1 Announce Type: new Abstract: Deep generative models provide state-of-the-art performance across a wide array of applications, with recent studies showing increasing applicability fo

Correct Answers from Sound Reasoning: Verifiable Process Supervision for Language Models

ResearchDGX agent

arXiv:2605.12519v1 Announce Type: cross Abstract: Training language models to produce both correct answers and sound reasoning remains an open challenge. Reinforcement learning with verifiable rewards

Do Androids Dream of Breaking the Game? Systematically Auditing AI Agent Benchmarks with BenchJack

Model ReleasesDGX agent

arXiv:2605.12673v1 Announce Type: new Abstract: Agent benchmarks have become the de facto measure of frontier AI competence, guiding model selection, investment, and deployment. However, reward hackin

Exploring how EFL students talk to and through AI to develop texts

ResearchDGX agent

arXiv:2605.12523v1 Announce Type: cross Abstract: Generative Artificial Intelligence (AI) introduces new considerations for English as a foreign language (EFL) writing pedagogy. This study explores ho

Generative Modeling from Black-box Corruptions via Self-Consistent Stochastic Interpolants

ResearchDGX agent

arXiv:2512.10857v2 Announce Type: replace-cross Abstract: Transport-based methods have emerged as a leading paradigm for building generative models from large, clean datasets. However, in many scienti

i recently left @Bloomberg after two great years. incredibly lucky to have joined @LangChain in time to work on this 🚀

AgentsDGX agent

i recently left @Bloomberg after two great years. incredibly lucky to have joined @LangChain in time to work on this 🚀 Spend less time on triaging Ship fixes faster Catch regressions earlier Introduci

Inference-Time Machine Unlearning via Gated Activation Redirection

Model ReleasesDGX agent

arXiv:2605.12765v1 Announce Type: new Abstract: Large Language Models memorize vast amounts of training data, raising concerns regarding privacy, copyright infringement, and safety. Machine unlearning

interwhen: A Generalizable Framework for Steering Reasoning Models with Test-time Verification

SafetyDGX agent

arXiv:2602.11202v3 Announce Type: replace-cross Abstract: Reasoning models produce long traces of intermediate decisions and tool calls, making test-time verification important for ensuring correctnes

KVServe: Service-Aware KV Cache Compression for Communication-Efficient Disaggregated LLM Serving

ApplicationsDGX agent

arXiv:2605.13734v1 Announce Type: cross Abstract: LLMs are widely adopted in production, pushing inference systems to their limits. Disaggregated LLM serving (e.g., PD separation and KV state disaggre

LeanSearch v2: Global Premise Retrieval for Lean 4 Theorem Proving

Model ReleasesDGX agent

arXiv:2605.13137v1 Announce Type: cross Abstract: Proving theorems in Lean 4 often requires identifying a scattered set of library lemmas whose joint use enables a concise proof -- a task we call glob

MobiBench: Multi-Branch, Modular Benchmark for Mobile GUI Agents

Model ReleasesDGX agent

arXiv:2512.12634v3 Announce Type: replace Abstract: Mobile GUI Agents, AI agents capable of interacting with mobile applications on behalf of users, have the potential to transform human computer inte

Protected Source Maps: Ship browser source maps securely

ToolsDGX agent

Protected Source Maps is a Vercel feature that enables developers to securely ship browser source maps to production without exposing them publicly. This allows for proper error tracking and debugging

Protocol-Driven Development: Governing Generated Software Through Invariants and Evidence

SafetyDGX agent

arXiv:2605.12981v1 Announce Type: cross Abstract: Automated program synthesis has reduced the cost of producing candidate implementations, but it introduces a harder governance problem: determining wh

Quoting Mitchell Hashimoto

AgentsDGX agent

[...] On the interesting side is how fungible programming languages are nowadays. Programming languages used to be LOCK IN, and they're increasingly not so. You think the Bun rewrite in Rust is good f

Qwen-Image-VAE-2.0 Technical Report

Model ReleasesDGX agent

arXiv:2605.13565v1 Announce Type: new Abstract: We present Qwen-Image-VAE-2.0, a suite of high-compression Variational Autoencoders (VAEs) that achieve significant advances in both reconstruction fide

Retrieval-Augmented Tutoring for Algorithm Tracing and Problem-Solving in AI Education

TutorialsDGX agent

arXiv:2605.12988v1 Announce Type: new Abstract: Students learning algorithms often need support as they interpret traces, debug reasoning errors, and apply procedures across unfamiliar problem instanc

Revealing Interpretable Failure Modes of VLMs

SafetyDGX agent

arXiv:2605.12674v1 Announce Type: new Abstract: Vision-Language Models (VLMs) are increasingly used in safety-critical applications because of their broad reasoning capabilities and ability to general

Revisiting DAgger in the Era of LLM-Agents

SafetyDGX agent

arXiv:2605.12913v1 Announce Type: new Abstract: Long-horizon LM agents learn from multi-turn interaction, where a single early mistake can alter the subsequent state distribution and derail the whole

Sample-Efficient Optimisation over the Outputs of Generative Models

ResearchDGX agent

arXiv:2509.23800v3 Announce Type: replace-cross Abstract: Modern generative AI models, such as diffusion and flow matching models, can sample from rich data distributions. However, many applications,

Sea's View on the Future of Agentic Software Development with Codex

AgentsDGX agent

This article presents Sea's perspective on how agentic software development—systems that can autonomously plan and execute coding tasks—will evolve with tools like Codex, OpenAI's code generation mode

SecurityScorecard acquires internet scanning startup Driftnet to bolster third-party risk platform

IndustryDGX agent

Cyber risk management company SecurityScorecard Inc. announced today that it has acquired Driftnet Ltd., a U.K.-based internet scanning and threat intelligence startup, in a deal aimed at bolting real

Seven papers. One research team. Together AI is heading to #MLSys2026 next week. Check out the work going from research to production on the…

ApplicationsDGX agent

Together AI will present seven research papers at MLSys 2026, showcasing projects that demonstrate the transition from research to production applications. The announcement highlights the company's co

Structural Diversity Drives Disruptive Scientific Innovation

SafetyDGX agent

arXiv:2605.12514v1 Announce Type: cross Abstract: Scientific innovation increasingly depends on collaboration, yet the organizational structure that fosters breakthrough ideas remains poorly understoo

The Readability Spectrum: Patterns, Issues, and Prompt Effects in LLM-Generated Code

ResearchDGX agent

arXiv:2605.13280v1 Announce Type: cross Abstract: As Large Language Models (LLMs) are transforming software development, the functional quality of generated code has become a central focus, leaving re

this is an early beta. we still have much to improve, which is why its so exciting for people to start trying. the goal is for Grok Build to…

AgentsDGX agent

this is an early beta. we still have much to improve, which is why its so exciting for people to start trying. the goal is for Grok Build to be the best coding agent when it comes to solving real and

Training Large Language Models to Predict Clinical Events

Model ReleasesDGX agent

arXiv:2605.12817v1 Announce Type: cross Abstract: Longitudinal clinical notes contain rich evidence of how patients evolve over time, but converting this signal into training supervision for clinical

When Diffusion Breaks Constraints: Sequential Autoregressive Generation with RL and MCTS

ResearchDGX agent

arXiv:2512.01242v3 Announce Type: replace-cross Abstract: Data-driven generative models excel in language and vision, but diffusion models often fail in constrained planning and design tasks, exhibiti

Yesterday we committed a cardinal sin: two first-party events in NYC, back-to-back. We had to close registration for both early. Packed room…

AgentsDGX agent

Yesterday we committed a cardinal sin: two first-party events in NYC, back-to-back. We had to close registration for both early. Packed rooms. Strong vibes. No regrets. 💻 Laptops out — developer works

Yesterday we hosted a wonderful happy hour 🍻 in NYC with @get_tabs. It was *packed* - we had 500+ signups, had to implement a waitlist, and…

ApplicationsDGX agent

Yesterday we hosted a wonderful happy hour 🍻 in NYC with @get_tabs. It was *packed* - we had 500+ signups, had to implement a waitlist, and the bar was full! Every attendee was an AI builder. It was a

Yield Curves Dynamics Using Variational Autoencoders Under No-arbitrage

ResearchDGX agent

arXiv:2605.12764v1 Announce Type: cross Abstract: This paper introduces a physics-informed generative framework that resolves the fundamental conflict between the statistical flexibility of deep learn

13 May 2026

1/5 We’re seeing 4 common agent optimization methods for hitting the right accuracy-cost or accuracy-latency tradeoff. We tried them all out…

AgentsDGX agent

AI21 Labs discusses four common agent optimization methods used to balance accuracy against cost and latency constraints. The post indicates the team evaluated all four approaches, likely covering tec

← Previous
1…7172737475…90
Next →