AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,598 results
Safety

Chain of Modality: From Static Fusion to Dynamic Orchestration in Omni-MLLMs

DGX agent

arXiv:2604.14520v1 Announce Type: new Abstract: Omni-modal Large Language Models (Omni-MLLMs) promise a unified integration of diverse sensory streams. However, recent evaluations reveal a critical pe

safetyarxiv-cs-cv
17 Apr 2026
Research
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Cognitive Alpha Mining via LLM-Driven Code-Based Evolution

DGX agent

arXiv:2511.18850v2 Announce Type: replace Abstract: Discovering effective predictive signals, or 'alphas,' from financial data with high dimensionality and extremely low signal-to-noise ratio remains

researcharxiv-cs-cl
17 Apr 2026
Model Releases

Contract-Coding: Towards Repo-Level Generation via Structured Symbolic Paradigm

DGX agent

arXiv:2604.13100v1 Announce Type: cross Abstract: The shift toward intent-driven software engineering (often termed 'Vibe Coding') exposes a critical Context-Fidelity Trade-off: vague user intents ove

model-releasesarxiv-cs-ai
17 Apr 2026
Research

Generative Augmented Inference

DGX agent

arXiv:2604.14575v1 Announce Type: new Abstract: Data-driven operations management often relies on parameters estimated from costly human-generated labels. Recent advances in large language models (LLM

researcharxiv-cs-lg
17 Apr 2026
Model Releases

HRDexDB: A Large-Scale Dataset of Dexterous Human and Robotic Hand Grasps

DGX agent

arXiv:2604.14944v1 Announce Type: cross Abstract: We present HRDexDB, a large-scale, multi-modal dataset of high-fidelity dexterous grasping sequences featuring both human and diverse robotic hands. U

model-releasesarxiv-cs-cv
17 Apr 2026
Tools

in retrospect putting the slop cannons (@_lopopolo) on @aiDotEngineer talks day 1 and putting the grown ups (@badlogicgames) on talks day 2 …

DGX agent

in retrospect putting the slop cannons (@_lopopolo) on @aiDotEngineer talks day 1 and putting the grown ups (@badlogicgames) on talks day 2 is working out pretty well for faithfully representing the m

toolsswyx--x
17 Apr 2026
Tools

In @steipete's latest State of the Claw, he gives an update on 5 months of @OpenClaw and some behind the scenes on what it's like maintainin…

DGX agent

In @steipete's latest State of the Claw, he gives an update on 5 months of @OpenClaw and some behind the scenes on what it's like maintaining the fastest growing open source of all time: https://www.y

toolsswyx--x
17 Apr 2026
Model Releases

MARCA: A Checklist-Based Benchmark for Multilingual Web Search

DGX agent

arXiv:2604.14448v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used as sources of information, yet their reliability depends on the ability to search the web, select rel

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

Mechanistic Decoding of Cognitive Constructs in LLMs

DGX agent

arXiv:2604.14593v1 Announce Type: new Abstract: While Large Language Models (LLMs) demonstrate increasingly sophisticated affective capabilities, the internal mechanisms by which they process complex

model-releasesarxiv-cs-cl
17 Apr 2026
Applications

Model-Based Reinforcement Learning under Random Observation Delays

DGX agent

arXiv:2509.20869v2 Announce Type: replace Abstract: Delays frequently occur in real-world environments, yet standard reinforcement learning (RL) algorithms often assume instantaneous perception of the

applicationsarxiv-cs-lg
17 Apr 2026
Hardware

Oh look! Anthropic's entire 'we are delaying Mythos' narrative was marketing hogwash. Kudos to FT for confirming what was obvious. Anthropic…

DGX agent

Oh look! Anthropic's entire 'we are delaying Mythos' narrative was marketing hogwash. Kudos to FT for confirming what was obvious. Anthropic simply doesn't have the compute. FT: 'Multiple people with

hardwaregary-marcus--x
17 Apr 2026
Safety

OmniCompliance-100K: A Multi-Domain, Rule-Grounded, Real-World Safety Compliance Dataset

DGX agent

arXiv:2603.13933v2 Announce Type: replace Abstract: Ensuring the safety and compliance of large language models (LLMs) is of paramount importance. However, existing LLM safety datasets often rely on a

safetyarxiv-cs-cl
17 Apr 2026
Model Releases

QuantCode-Bench: A Benchmark for Evaluating the Ability of Large Language Models to Generate Executable Algorithmic Trading Strategies

DGX agent

arXiv:2604.15151v1 Announce Type: new Abstract: Large language models have demonstrated strong performance on general-purpose programming tasks, yet their ability to generate executable algorithmic tr

model-releasesarxiv-cs-cl
17 Apr 2026
Safety

RECOVER: Designing a Large Language Model-based Remote Patient Monitoring System for Postoperative Gastrointestinal Cancer Care

DGX agent

arXiv:2502.05740v2 Announce Type: replace-cross Abstract: Cancer surgery is a key treatment for gastrointestinal (GI) cancers, a group of cancers that account for more than 35% of cancer-related death

safetyarxiv-cs-ai
17 Apr 2026
Safety

SPAGBias: Uncovering and Tracing Structured Spatial Gender Bias in Large Language Models

DGX agent

arXiv:2604.14672v1 Announce Type: new Abstract: Large language models (LLMs) are being increasingly used in urban planning, but since gendered space theory highlights how gender hierarchies are embedd

safetyarxiv-cs-cl
17 Apr 2026
Safety

The Autocorrelation Blind Spot: Why 42% of Turn-Level Findings in LLM Conversation Analysis May Be Spurious

DGX agent

arXiv:2604.14414v1 Announce Type: new Abstract: Turn-level metrics are widely used to evaluate properties of multi-turn human-LLM conversations, from safety and sycophancy to dialogue quality. However

safetyarxiv-cs-cl
17 Apr 2026
Industry

The Inference Cloud Memory Layer: A Technical Dive into DigitalOcean Managed Databases

DGX agent

DigitalOcean's Inference Cloud Memory Layer is a technical architecture component designed to optimize database performance by implementing an in-memory caching layer for faster data access and reduce

industrydigitalocean
17 Apr 2026
Safety

The PICCO Framework for Large Language Model Prompting: A Taxonomy and Reference Architecture for Prompt Structure

DGX agent

arXiv:2604.14197v1 Announce Type: new Abstract: Large language model (LLM) performance depends heavily on prompt design, yet prompt construction is often described and applied inconsistently. Our purp

safetyarxiv-cs-cl
17 Apr 2026
Safety

VoxSafeBench: Not Just What Is Said, but Who, How, and Where

DGX agent

arXiv:2604.14548v1 Announce Type: cross Abstract: As speech language models (SLMs) transition from personal devices into shared, multi-user environments, their responses must account for far more than

safetyarxiv-cs-lg
17 Apr 2026
Applications

Abstract 3D Perception for Spatial Intelligence in Vision-Language Models

DGX agent

arXiv:2511.10946v3 Announce Type: replace Abstract: Vision-language models (VLMs) struggle with 3D-related tasks such as spatial cognition and physical understanding, which are crucial for real-world

applicationsarxiv-cs-cv
16 Apr 2026
Safety

Activation-Guided Local Editing for Jailbreaking Attacks

DGX agent

arXiv:2508.00555v2 Announce Type: replace-cross Abstract: Jailbreaking is an essential adversarial technique for red-teaming these models to uncover and patch security flaws. However, existing jailbre

safetyarxiv-cs-cl
16 Apr 2026
Safety

Alignment as Institutional Design: From Behavioral Correction to Transaction Structure in Intelligent Systems

DGX agent

arXiv:2604.13079v1 Announce Type: cross Abstract: Current AI alignment paradigms rely on behavioral correction: external supervisors (e.g., RLHF) observe outputs, judge against preferences, and adjust

safetyarxiv-cs-lg
16 Apr 2026
Model Releases

also available on the Claude Blog: https://claude.com/blog/using-claude-code-session-management-and-1m-context

DGX agent

Claude Code's session management capabilities and 1 million token context window are highlighted in this post, which references an official Anthropic blog entry. The feature allows developers to maint

model-releasesthariq--x
16 Apr 2026
Research

Automated co-design of high-performance thermodynamic cycles via graph-based hierarchical reinforcement learning

DGX agent

arXiv:2604.13133v1 Announce Type: new Abstract: Thermodynamic cycles are pivotal in determining the efficacy of energy conversion systems. Traditional design methodologies, which rely on expert knowle

researcharxiv-cs-lg
16 Apr 2026
Safety

Character Beyond Speech: Leveraging Role-Playing Evaluation in Audio Large Language Models via Reinforcement Learning

DGX agent

arXiv:2604.13804v1 Announce Type: new Abstract: The rapid evolution of multimodal large models has revolutionized the simulation of diverse characters in speech dialogue systems, enabling a novel inte

safetyarxiv-cs-lg
16 Apr 2026
Model Releases

Databricks on Google Cloud: Innovate Faster. Smarter. Together.

DGX agent

Databricks and Google Cloud have partnered to enable organizations to build and deploy data and AI solutions more efficiently. The collaboration integrates Databricks' lakehouse platform with Google C

model-releasesdatabricks
16 Apr 2026
Research

Design Conditions for Intra-Group Learning of Sequence-Level Rewards: Token Gradient Cancellation

DGX agent

arXiv:2604.13088v1 Announce Type: new Abstract: In sparse termination rewards, intra-group comparisons have become the dominant paradigm for fine-tuning reasoning models via reinforcement learning. Ho

researcharxiv-cs-lg
16 Apr 2026
Applications

Designing synthetic datasets for the real world: Mechanism design and reasoning from first principles

DGX agent

This Google Research work presents guidelines for synthetic data mechanism design and provides insights into generating and evaluating synthetic data at scale. The research introduces a reasoning-driv

applicationsgoogle-research
16 Apr 2026
Industry

Developer tooling startup Expo nabs $45M investment

DGX agent

Expo, the developer of a popular open-source tool for building cross-platform applications, today announced that it has raised 45 million in funding. Developers often implement web application interfa

industrysiliconangle
16 Apr 2026
Research

Dual-Enhancement Product Bundling: Bridging Interactive Graph and Large Language Model

DGX agent

arXiv:2604.14030v1 Announce Type: new Abstract: Product bundling boosts e-commerce revenue by recommending complementary item combinations. However, existing methods face two critical challenges: (1)

researcharxiv-cs-cl
16 Apr 2026
Model Releases

EmbodiedClaw: Conversational Workflow Execution for Embodied AI Development

DGX agent

arXiv:2604.13800v1 Announce Type: new Abstract: Embodied AI research is increasingly moving beyond single-task, single-environment policy learning toward multi-task, multi-scene, and multi-model setti

model-releasesarxiv-cs-ro
16 Apr 2026
Model Releases

ESCAPE: Episodic Spatial Memory and Adaptive Execution Policy for Long-Horizon Mobile Manipulation

DGX agent

arXiv:2604.13633v1 Announce Type: new Abstract: Coordinating navigation and manipulation with robust performance is essential for embodied AI in complex indoor environments. However, as tasks extend o

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

I edited the intro because I realized I buried the lede originally- The 1M context window is a double-edged sword. It allows Claude to do mo…

DGX agent

I edited the intro because I realized I buried the lede originally- The 1M context window is a double-edged sword. It allows Claude to do more complex tasks but it can also leads to more context pollu

model-releasesthariq--x
16 Apr 2026
Tools

I love to explore solutions with it so I often ask it to brainstorm with me, then choose an option and rewind to implement it. When I'm read…

DGX agent

I love to explore solutions with it so I often ask it to brainstorm with me, then choose an option and rewind to implement it. When I'm ready, I ask it to interview me to figure out what’s in my head

toolsthariq--x
16 Apr 2026
Model Releases

IndicDB -- Benchmarking Multilingual Text-to-SQL Capabilities in Indian Languages

DGX agent

arXiv:2604.13686v1 Announce Type: new Abstract: While Large Language Models (LLMs) have significantly advanced Text-to-SQL performance, existing benchmarks predominantly focus on Western contexts and

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

LaoBench: A Large-Scale Multidimensional Lao Benchmark for Large Language Models

DGX agent

arXiv:2511.11334v3 Announce Type: replace Abstract: The rapid advancement of large language models (LLMs) has not been matched by their evaluation in low-resource languages, especially Southeast Asian

model-releasesarxiv-cs-cl
16 Apr 2026
Tutorials

Learn more about our journey https://youtu.be/aaBRSWWB_tI

DGX agent

Replit shared a YouTube video detailing their company's journey and history of development. The video likely covers key milestones, founding story, product evolution, and the team's vision for their c

tutorialsreplit--x
16 Apr 2026
Model Releases

Lossless Prompt Compression via Dictionary-Encoding and In-Context Learning: Enabling Cost-Effective LLM Analysis of Repetitive Data

DGX agent

arXiv:2604.13066v1 Announce Type: new Abstract: In-context learning has established itself as an important learning paradigm for Large Language Models (LLMs). In this paper, we demonstrate that LLMs c

model-releasesarxiv-cs-cl
16 Apr 2026
Safety

Multi-Dimensional Knowledge Profiling with Large-Scale Literature Database and Hierarchical Retrieval

DGX agent

arXiv:2601.15170v2 Announce Type: replace Abstract: The rapid expansion of research across machine learning, vision, and language has produced a volume of publications that is increasingly difficult t

safetyarxiv-cs-cv
16 Apr 2026
Model Releases

Olfactory pursuit: catching a moving odor source in complex flows

DGX agent

arXiv:2604.13121v1 Announce Type: new Abstract: Locating and intercepting a moving target from possibly delayed, intermittent sensory signals is a paradigmatic problem in decision-making under uncerta

model-releasesarxiv-cs-ro
16 Apr 2026
Model Releases

🎬 Ollama Gemma Day Recap: SGLang at the Ollama Gemma 4 Party in Palo Alto 🍾 Last night, @ollama hosted a packed Gemma Day at the Palo Alto…

DGX agent

🎬 Ollama Gemma Day Recap: SGLang at the Ollama Gemma 4 Party in Palo Alto 🍾 Last night, @ollama hosted a packed Gemma Day at the Palo Alto office alongside the @GoogleDeepMind Gemma team. SGLang was i

model-releasesollama--x
16 Apr 2026
Model Releases

Peer-Predictive Self-Training for Language Model Reasoning

DGX agent

arXiv:2604.13356v1 Announce Type: new Abstract: Mechanisms for continued self-improvement of language models without external supervision remain an open challenge. We propose Peer-Predictive Self-Trai

model-releasesarxiv-cs-cl
16 Apr 2026
Local Ai

qwen3.6 is out

DGX agent

Qwen 3.6 Plus Preview is Alibaba's next-generation large language model released on March 30-31, 2026 , and the first open-weight variant was released following the February Qwen 3.5 series, prioritiz

local-air-ollama
16 Apr 2026
Model Releases

Reward Hacking in the Era of Large Models: Mechanisms, Emergent Misalignment, Challenges

DGX agent

arXiv:2604.13602v1 Announce Type: new Abstract: Reinforcement Learning from Human Feedback (RLHF) and related alignment paradigms have become central to steering large language models (LLMs) and multi

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

RPS: Information Elicitation with Reinforcement Prompt Selection

DGX agent

arXiv:2604.13817v1 Announce Type: new Abstract: Large language models (LLMs) have shown remarkable capabilities in dialogue generation and reasoning, yet their effectiveness in eliciting user-known bu

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Seek-and-Solve: Benchmarking MLLMs for Visual Clue-Driven Reasoning in Daily Scenarios

DGX agent

arXiv:2604.14041v1 Announce Type: new Abstract: Daily scenarios are characterized by visual richness, requiring Multimodal Large Language Models (MLLMs) to filter noise and identify decisive visual cl

model-releasesarxiv-cs-cv
16 Apr 2026
Safety

Self-adaptive Multi-Access Edge Architectures: A Robotics Case

DGX agent

arXiv:2604.13542v1 Announce Type: new Abstract: The growth of compute-intensive AI tasks highlights the need to mitigate the processing costs and improve performance and energy efficiency. This necess

safetyarxiv-cs-ro
16 Apr 2026
Model Releases

Shocking result on my pelican benchmark this morning, I got a better pelican from a 21GB local Qwen3.6-35B-A3B running on my laptop than I d…

DGX agent

Shocking result on my pelican benchmark this morning, I got a better pelican from a 21GB local Qwen3.6-35B-A3B running on my laptop than I did from the new Opus 4.7! Qwen on the left, Opus on the righ

model-releasesclem-delangue--x
16 Apr 2026
← Previous
1…359360361362363…367
Next →