AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,424 results
Agents

TableNet A Large-Scale Table Dataset with LLM-Powered Autonomous

DGX agent

arXiv:2604.13041v1 Announce Type: cross Abstract: Table Structure Recognition (TSR) requires the logical reasoning ability of large language models (LLMs) to handle complex table layouts, but current

agentsarxiv-cs-ai
17 Apr 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

The LLM Fallacy: Misattribution in AI-Assisted Cognitive Workflows

DGX agent

arXiv:2604.14807v1 Announce Type: cross Abstract: The rapid integration of large language models (LLMs) into everyday workflows has transformed how individuals perform cognitive tasks such as writing,

safetyarxiv-cs-cl
17 Apr 2026
Safety

The PICCO Framework for Large Language Model Prompting: A Taxonomy and Reference Architecture for Prompt Structure

DGX agent

arXiv:2604.14197v1 Announce Type: new Abstract: Large language model (LLM) performance depends heavily on prompt design, yet prompt construction is often described and applied inconsistently. Our purp

safetyarxiv-cs-cl
17 Apr 2026
Model Releases

Time-RA: Towards Time Series Reasoning for Anomaly Diagnosis with LLM Feedback

DGX agent

arXiv:2507.15066v5 Announce Type: replace Abstract: Time series anomaly detection (TSAD) has traditionally focused on binary classification and often lacks the fine-grained categorization and explanat

model-releasesarxiv-cs-lg
17 Apr 2026
Model Releases

What folk don’t get is that the play of Anthropic et al is to get enterprises to use their models via their own wrapper hooked into systems …

DGX agent

What folk don’t get is that the play of Anthropic et al is to get enterprises to use their models via their own wrapper hooked into systems of record This is why they are moving from per seat pricing

model-releasesemad-mostaque--x
17 Apr 2026
Applications

You can watch the accelerated shipping from the AI labs to get a feeling of what AI-driven product development makes possible. A tremendous …

DGX agent

You can watch the accelerated shipping from the AI labs to get a feeling of what AI-driven product development makes possible. A tremendous number of products are coming out, many of them are really g

applicationsethan-mollick--x
17 Apr 2026
Safety

A Mechanistic Analysis of Sim-and-Real Co-Training in Generative Robot Policies

DGX agent

arXiv:2604.13645v1 Announce Type: cross Abstract: Co-training, which combines limited in-domain real-world data with abundant surrogate data such as simulation or cross-embodiment robot data, is widel

safetyarxiv-cs-lg
16 Apr 2026
Model Releases

A Study of Failure Modes in Two-Stage Human-Object Interaction Detection

DGX agent

arXiv:2604.13448v1 Announce Type: new Abstract: Human-object interaction (HOI) detection aims to detect interactions between humans and objects in images. While recent advances have improved performan

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

Accelerating the cyber defense ecosystem that protects us all

DGX agent

OpenAI discusses efforts to strengthen collective cybersecurity defenses through collaboration and technology advancement within the broader cyber defense ecosystem. The article likely covers OpenAI's

model-releasesopenai
16 Apr 2026
Model Releases

Adaptive Multi-Scale Channel-Spatial Attention Aggregation Framework for 3D Indoor Semantic Scene Completion Toward Assisting Visually Impaired

DGX agent

arXiv:2602.16385v4 Announce Type: replace Abstract: Independent indoor mobility remains a critical challenge for individuals with visual impairments, largely due to the limited capability of existing

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

AI Powered Image Analysis for Phishing Detection

DGX agent

arXiv:2604.13555v1 Announce Type: new Abstract: Phishing websites now rely heavily on visual imitation-copied logos, similar layouts, and matching colours-to avoid detection by text- and URL-based sys

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

An Optimal Transport-driven Approach for Cultivating Latent Space in Online Incremental Learning

DGX agent

arXiv:2211.16780v3 Announce Type: replace-cross Abstract: In online incremental learning, data continuously arrives with substantial distributional shifts, creating a significant challenge because pre

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

Auto-FP: An Experimental Study of Automated Feature Preprocessing for Tabular Data

DGX agent

arXiv:2310.02540v2 Announce Type: replace Abstract: Classical machine learning models, such as linear models and tree-based models, are widely used in industry. These models are sensitive to data dist

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Bridging MARL to SARL: An Order-Independent Multi-Agent Transformer via Latent Consensus

DGX agent

arXiv:2604.13472v1 Announce Type: new Abstract: Cooperative multi-agent reinforcement learning (MARL) is widely used to address large joint observation and action spaces by decomposing a centralized c

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Can Large Language Models Reliably Extract Physiology Index Values from Coronary Angiography Reports?

DGX agent

arXiv:2604.13077v1 Announce Type: new Abstract: Coronary angiography (CAG) reports contain clinically relevant physiological measurements, yet this information is typically in the form of unstructured

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

CodeFlowBench: A Multi-turn, Iterative Benchmark for Complex Code Generation

DGX agent

arXiv:2504.21751v4 Announce Type: replace-cross Abstract: Modern software development demands code that is maintainable, testable, and scalable by organizing the implementation into modular components

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

ExpSeek: Self-Triggered Experience Seeking for Web Agents

DGX agent

arXiv:2601.08605v2 Announce Type: replace Abstract: Experience intervention in web agents emerges as a promising technical paradigm, enhancing agent interaction capabilities by providing valuable insi

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

F-Actor: Controllable Conversational Behaviour in Full-Duplex Models

DGX agent

arXiv:2601.11329v3 Announce Type: replace Abstract: Spoken conversational systems require more than accurate speech generation to have human-like conversations: to feel natural and engaging, they must

model-releasesarxiv-cs-cl
16 Apr 2026
Safety

From Alignment to Prediction: A Study of Self-Supervised Learning and Predictive Representation Learning

DGX agent

arXiv:2604.13518v1 Announce Type: new Abstract: Self-supervised learning has emerged as a major technique for the task of learning from unlabeled data, where the current methods mostly revolve around

safetyarxiv-cs-lg
16 Apr 2026
Model Releases

Here's Qwen 3.6-35B-A3B v.s. Claude Opus 4.7 for 'Generate an SVG of a flamingo riding a unicycle', in case you thought Qwen might be cheati…

DGX agent

This post compares the performance of Qwen 3.6-35B-A3B and Claude Opus 4.7 models on a creative task of generating SVG code for a flamingo riding a unicycle, likely demonstrating differences in their

model-releasessimon-willison--x
16 Apr 2026
Model Releases

HOLY 🤯 The one and only @elder_plinius just dropped an unlocked Gemma 4 E4B, and the specs are INSANE. Look at the performance shifts: → Re…

DGX agent

HOLY 🤯 The one and only @elder_plinius just dropped an unlocked Gemma 4 E4B, and the specs are INSANE. Look at the performance shifts: → Refusal rate: 98.8% down to 2.1% (!!) → Compliance: 1.2% up to

model-releasesclem-delangue--x
16 Apr 2026
Model Releases

How WPP accelerates humanoid robot training 10x with G4 VMs

DGX agent

Editor’s note: Today we hear from Perry Nightingale, SVP of Creative AI at WPP about the workflow that cuts training time for humanoid robots from days to minutes — plus access to the open-source code

model-releasesgoogle-cloud-ai
16 Apr 2026
Tutorials

I've got a lot to say here! I'll post a guide on how to prompt with Opus 4.7 as well as talk about a personal project I've been working on u…

DGX agent

I've got a lot to say here! I'll post a guide on how to prompt with Opus 4.7 as well as talk about a personal project I've been working on using it. I hope you enjoy working with Opus 4.7 and getting

tutorialsthariq--x
16 Apr 2026
Model Releases

LaoBench: A Large-Scale Multidimensional Lao Benchmark for Large Language Models

DGX agent

arXiv:2511.11334v3 Announce Type: replace Abstract: The rapid advancement of large language models (LLMs) has not been matched by their evaluation in low-resource languages, especially Southeast Asian

model-releasesarxiv-cs-cl
16 Apr 2026
Agents

LEO-RobotAgent: A General-purpose Robotic Agent for Language-driven Embodied Operator

DGX agent

arXiv:2512.10605v2 Announce Type: replace Abstract: We propose LEO-RobotAgent, a general-purpose language-driven intelligent agent framework for robots. Under this framework, LLMs can operate differen

agentsarxiv-cs-ro
16 Apr 2026
Applications

Lite Any Stereo: Efficient Zero-Shot Stereo Matching

DGX agent

arXiv:2511.16555v3 Announce Type: replace Abstract: Recent advances in stereo matching have focused on accuracy, often at the cost of significantly increased model size. Traditionally, the community h

applicationsarxiv-cs-cv
16 Apr 2026
Local Ai

mcpstrike – Let your local LLM run the pentest for you

DGX agent

**mcpstrike** is a tool shared on the r/ollama subreddit that integrates a locally-run LLM (via Ollama) with an MCP (Model Context Protocol) server to autonomously execute penetration testing workflow

local-air-ollama
16 Apr 2026
Model Releases

MERRIN: A Benchmark for Multimodal Evidence Retrieval and Reasoning in Noisy Web Environments

DGX agent

arXiv:2604.13418v1 Announce Type: new Abstract: Motivated by the underspecified, multi-hop nature of search queries and the multimodal, heterogeneous, and often conflicting nature of real-world web re

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Mozilla launches Thunderbolt AI client with focus on self-hosted infrastructure

DGX agent

Thunderbolt is a new open-source AI client from Mozilla-owned MZLA Technologies aimed at enterprises who want to run self-hosted chatbots on their own infrastructure. The platform allows organizations

model-releasesars-technica
16 Apr 2026
Model Releases

MulDimIF: A Multi-Dimensional Constraint Framework for Evaluating and Improving Instruction Following in Large Language Models

DGX agent

arXiv:2505.07591v2 Announce Type: replace Abstract: Instruction following refers to the ability of large language models (LLMs) to generate outputs that satisfy all specified constraints. Existing res

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Multi-Task LLM with LoRA Fine-Tuning for Automated Cancer Staging and Biomarker Extraction

DGX agent

arXiv:2604.13328v1 Announce Type: new Abstract: Pathology reports serve as the definitive record for breast cancer staging, yet their unstructured format impedes large-scale data curation. While Large

model-releasesarxiv-cs-lg
16 Apr 2026
Industry

OpenAI starts offering a biology-tuned LLM

DGX agent

OpenAI has launched GPT-Rosalind, a biology-tuned large language model . This specialized LLM is designed to enhance performance on biology-specific tasks and applications. The model represents OpenAI

industryars-technica
16 Apr 2026
Industry

OpenAI updates its Codex desktop app with features like computer control, an in-app browser, image generation, automation memory, plugin support, and more (David Gewirtz/ZDNET)

DGX agent

David Gewirtz / ZDNET: OpenAI updates its Codex desktop app with features like computer control, an in-app browser, image generation, automation memory, plugin support, and more — ZDNET's key takeaway

industrytechmeme
16 Apr 2026
Model Releases

OPTED: Open Preprocessed Trachoma Eye Dataset Using Zero-Shot SAM 3 Segmentation

DGX agent

arXiv:2603.06885v2 Announce Type: replace Abstract: Trachoma remains the leading infectious cause of blindness worldwide, with Sub-Saharan Africa bearing over 85% of the global burden and Ethiopia alo

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

Out of Context: Reliability in Multimodal Anomaly Detection Requires Contextual Inference

DGX agent

arXiv:2604.13252v1 Announce Type: new Abstract: Anomaly detection aims to identify observations that deviate from expected behavior. Because anomalous events are inherently sparse, most frameworks are

model-releasesarxiv-cs-lg
16 Apr 2026
Agents

Remember Me, Refine Me: A Dynamic Procedural Memory Framework for Experience-Driven Agent Evolution

DGX agent

arXiv:2512.10696v2 Announce Type: replace-cross Abstract: Procedural memory enables large language model (LLM) agents to internalize 'how-to' knowledge, theoretically reducing redundant trial-and-erro

agentsarxiv-cs-cl
16 Apr 2026
Safety

Representation over Routing: Overcoming Surrogate Hacking in Multi-Timescale PPO

DGX agent

arXiv:2604.13517v1 Announce Type: new Abstract: Temporal credit assignment in reinforcement learning has long been a central challenge. Inspired by the multi-timescale encoding of the dopamine system

safetyarxiv-cs-lg
16 Apr 2026
Safety

RFK Jr. forces FDA to reconsider 12 unproven peptides after 2023 ban

DGX agent

In 2023, the FDA removed 19 peptides from the list of drugs that compounding pharmacies could produce, and in 2026 the FDA announced it will review whether to add back 7 of these peptides following pr

safetyars-technica
16 Apr 2026
Safety

Soft Q(lambda): A multi-step off-policy method for entropy regularised reinforcement learning using eligibility traces

DGX agent

arXiv:2604.13780v1 Announce Type: new Abstract: Soft Q-learning has emerged as a versatile model-free method for entropy-regularised reinforcement learning, optimising for returns augmented with a pen

safetyarxiv-cs-lg
16 Apr 2026
Safety

Spectral methods: crucial for machine learning, natural for quantum computers?

DGX agent

arXiv:2603.24654v2 Announce Type: replace-cross Abstract: This article presents an argument for why quantum computers could unlock new methods for machine learning. We argue that spectral methods, in

safetyarxiv-cs-lg
16 Apr 2026
Agents

this is fun. been trying eeg headsets for 10 yrs, the hat makes so much sense :)

DGX agent

this is fun. been trying eeg headsets for 10 yrs, the hat makes so much sense :) you can now control things with your brain. literally. we're building the most wearable BCI on the planet, with @sabica

agentsyohei-nakajima--x
16 Apr 2026
Industry

TSMC CEO C.C. Wei says TSMC checked with customers about AI demand and was reassured that it was still strong amid the Iran war, as it raises revenue forecasts (Wall Street Journal)

DGX agent

Wall Street Journal: TSMC CEO C.C. Wei says TSMC checked with customers about AI demand and was reassured that it was still strong amid the Iran war, as it raises revenue forecasts — Taiwan company ex

industrytechmeme
16 Apr 2026
Safety

UMI-3D: Extending Universal Manipulation Interface from Vision-Limited to 3D Spatial Perception

DGX agent

arXiv:2604.14089v1 Announce Type: new Abstract: We present UMI-3D, a multimodal extension of the Universal Manipulation Interface (UMI) for robust and scalable data collection in embodied manipulation

safetyarxiv-cs-ro
16 Apr 2026
Safety

Why dynamically routing multi-timescale advantages in PPO causes policy collapse (and a simple decoupled fix) [R]

DGX agent

This Reddit post discusses a known instability in PPO when advantage estimates operating across different temporal scales (e.g., short-horizon and long-horizon returns) are dynamically routed or mixed

safetyr-machinelearning
16 Apr 2026
Applications

390M+ embeddings. 100K+ namespaces. Sustained P50 latency of ~60ms at ~40 QPS. 📈 @ZoomInfo used our Dedicated Read Nodes, now in GA, to byp…

DGX agent

390M+ embeddings. 100K+ namespaces. Sustained P50 latency of ~60ms at ~40 QPS. 📈 @ZoomInfo used our Dedicated Read Nodes, now in GA, to bypass the 'infrastructure wall' and ship real-time AI recommend

applicationspinecone--x
15 Apr 2026
Safety

A Comparison of Reinforcement Learning and Optimal Control Methods for Path Planning

DGX agent

arXiv:2604.12628v1 Announce Type: cross Abstract: Path-planning for autonomous vehicles in threat-laden environments is a fundamental challenge. While traditional optimal control methods can find idea

safetyarxiv-cs-ro
15 Apr 2026
Safety

A longitudinal health agent framework

DGX agent

arXiv:2604.12019v1 Announce Type: new Abstract: Although artificial intelligence (AI) agents are increasingly proposed to support potentially longitudinal health tasks, such as symptom management, beh

safetyarxiv-cs-ai
15 Apr 2026
Agents

Agentic Discovery with Active Hypothesis Exploration for Visual Recognition

DGX agent

arXiv:2604.12999v1 Announce Type: new Abstract: We introduce HypoExplore, an agentic framework that formulates neural architecture discovery for visual recognition as a hypothesis-driven scientific in

agentsarxiv-cs-cv
15 Apr 2026
← Previous
1…521522523524525…530
Next →