AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,629 results
14 Apr 2026

SciPredict: Can LLMs Predict the Outcomes of Scientific Experiments in Natural Sciences?

Model ReleasesDGX agent

arXiv:2604.10718v1 Announce Type: new Abstract: Accelerating scientific discovery requires the identification of which experiments would yield the best outcomes before committing resources to costly p

SemaClaw: A Step Towards General-Purpose Personal AI Agents through Harness Engineering

SafetyDGX agent

arXiv:2604.11548v1 Announce Type: new Abstract: The rise of OpenClaw in early 2026 marks the moment when millions of users began deploying personal AI agents into their daily lives, delegating tasks r

Simulating Organized Group Behavior: New Framework, Benchmark, and Analysis

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.09874v1 Announce Type: new Abstract: Simulating how organized groups (e.g., corporations) make decisions (e.g., responding to a competitor's move) is essential for understanding real-world

SpatialScore: Towards Comprehensive Evaluation for Spatial Intelligence

Model ReleasesDGX agent

arXiv:2505.17012v3 Announce Type: replace-cross Abstract: Existing evaluations of multimodal large language models (MLLMs) on spatial intelligence are typically fragmented and limited in scope. In thi

SRBench: A Comprehensive Benchmark for Sequential Recommendation with Large Language Models

Model ReleasesDGX agent

arXiv:2604.09553v1 Announce Type: cross Abstract: LLM development has aroused great interest in Sequential Recommendation (SR) applications. However, comprehensive evaluation of SR models remains lack

StarVLA-alpha: Reducing Complexity in Vision-Language-Action Systems

Model ReleasesDGX agent

arXiv:2604.11757v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have recently emerged as a promising paradigm for building general-purpose robotic agents. However, the VLA landsc

Steered LLM Activations are Non-Surjective

SafetyDGX agent

arXiv:2604.09839v1 Announce Type: new Abstract: Activation steering is a popular white-box control technique that modifies model activations to elicit an abstract change in output behavior. It has als

The Blind Spot of Agent Safety: How Benign User Instructions Expose Critical Vulnerabilities in Computer-Use Agents

Model ReleasesDGX agent

arXiv:2604.10577v1 Announce Type: cross Abstract: Computer-use agents (CUAs) can now autonomously complete complex tasks in real digital environments, but when misled, they can also be used to automat

The @MiniMax_AI team is receptive to community feedback and are updating the license. See the latest here. https://x.com/RyanLeeMiniMax/stat…

Local AiDGX agent

The @MiniMax_AI team is receptive to community feedback and are updating the license. See the latest here. https://x.com/RyanLeeMiniMax/status/2044132777877221515?s=20 I just updated our license. For

The Paradox of Professional Input: How Expert Collaboration with AI Systems Shapes Their Future Value

SafetyDGX agent

arXiv:2504.12654v1 Announce Type: cross Abstract: This perspective paper examines a fundamental paradox in the relationship between professional expertise and artificial intelligence: as domain expert

The second technique was two-phase post-training. We first trained purely for capability, then added a latency penalty calibrated from real …

TutorialsDGX agent

The second technique was two-phase post-training. We first trained purely for capability, then added a latency penalty calibrated from real dogfooding data based on the CDF of how long users stay on S

There are multiple realities of AI right now. And what you have access to drastically changes your workflows, trust in AI, and ability to ad…

Model ReleasesDGX agent

There are multiple realities of AI right now. And what you have access to drastically changes your workflows, trust in AI, and ability to adapt to the future. Here’s the briefest state of the AI world

(This is a big part of what was called emergence in earlier academic work on unexpected LLM ability gains)

ApplicationsDGX agent

Ethan Mollick discusses the concept of 'emergence' in large language models (LLMs), referring to the phenomenon where AI systems appear to suddenly develop unexpected capabilities as they scale. The p

Today is the launch of 'WE ARE AS GODS: A Survival Guide for the Age of Abundance' -- my best book ever. Please check it out! The follow-up …

TutorialsDGX agent

Peter Diamandis announced the launch of his book 'WE ARE AS GODS: A Survival Guide for the Age of Abundance,' which he described as his best work to date. The book likely explores themes of technologi

Towards Autonomous Mechanistic Reasoning in Virtual Cells

AgentsDGX agent

arXiv:2604.11661v1 Announce Type: cross Abstract: Large language models (LLMs) have recently gained significant attention as a promising approach to accelerate scientific discovery. However, their app

UniToolCall: Unifying Tool-Use Representation, Data, and Evaluation for LLM Agents

Model ReleasesDGX agent

arXiv:2604.11557v1 Announce Type: new Abstract: Tool-use capability is a fundamental component of LLM agents, enabling them to interact with external systems through structured function calls. However

US affidavit: the man charged with attacking Sam Altman's home had a document that 'identified views opposed' to AI and listed addresses of other AI executives (New York Times)

IndustryDGX agent

New York Times: US affidavit: the man charged with attacking Sam Altman's home had a document that “identified views opposed” to AI and listed addresses of other AI executives — The authorities said a

Variance-Aware Prior-Based Tree Policies for Monte Carlo Tree Search

SafetyDGX agent

arXiv:2512.21648v2 Announce Type: replace-cross Abstract: Monte Carlo Tree Search (MCTS) has profoundly influenced reinforcement learning (RL) by integrating planning and learning in tasks requiring l

VGA-Bench: A Unified Benchmark and Multi-Model Framework for Video Aesthetics and Generation Quality Evaluation

Model ReleasesDGX agent

arXiv:2604.10127v1 Announce Type: cross Abstract: The rapid advancement of AIGC-based video generation has underscored the critical need for comprehensive evaluation frameworks that go beyond traditio

Virtual Smart Metering in District Heating Networks via Heterogeneous Spatial-Temporal Graph Neural Networks

Model ReleasesDGX agent

arXiv:2604.10166v1 Announce Type: cross Abstract: Intelligent operation of thermal energy networks aims to improve energy efficiency, reliability, and operational flexibility through data-driven contr

VS-Bench: Evaluating VLMs for Strategic Abilities in Multi-Agent Environments

Model ReleasesDGX agent

arXiv:2506.02387v3 Announce Type: replace Abstract: Recent advancements in Vision Language Models (VLMs) have expanded their capabilities to interactive agent tasks, yet existing benchmarks remain lim

Who's going to be at AIE Miami! I wanna hang with you.

ToolsDGX agent

This appears to be a social media post from Swyx (likely Shawn Wang, a developer advocate and AI enthusiast) on X (formerly Twitter), seeking to connect with other attendees at an AI Engineer (AIE) ev

YIELD: A Large-Scale Dataset and Evaluation Framework for Information Elicitation Agents

SafetyDGX agent

arXiv:2604.10968v1 Announce Type: new Abstract: Most conversational agents (CAs) are designed to satisfy user needs through user-driven interactions. However, many real-world settings, such as academi

13 Apr 2026

$200/month is enough to buy an H100 GPU for 6 hours every workday

HardwareDGX agent

Soumith Chintala shared a post highlighting that $200 per month is sufficient to rent access to an NVIDIA H100 GPU for approximately 6 hours every workday, making high-end AI compute more accessible t

Adaptive Simulation Experiment for LLM Policy Optimization

Model ReleasesDGX agent

arXiv:2604.08779v1 Announce Type: new Abstract: Large language models (LLMs) have significant potential to improve operational efficiency in operations management. Deploying these models requires spec

AI chatbots misdiagnose in over 80% of early medical cases, study finds https://ft.trib.al/cRifiz2

SafetyDGX agent

A study has found that AI chatbots incorrectly diagnose patients in more than 80% of early medical cases, raising significant concerns about the reliability of AI tools in clinical settings. The resea

AI-Induced Human Responsibility (AIHR) in AI-Human teams

AgentsDGX agent

arXiv:2604.08866v1 Announce Type: cross Abstract: As organizations increasingly deploy AI as a teammate rather than a standalone tool, morally consequential mistakes often arise from joint human-AI wo

As AI agents accelerate coding, what is the future of software engineering? Some trends are clear, such as the Product Management Bottleneck…

SafetyDGX agent

As AI agents accelerate coding, what is the future of software engineering? Some trends are clear, such as the Product Management Bottleneck, referring to the idea that we are more constrained by deci

Automated Batch Distillation Process Simulation for a Large Hybrid Dataset for Deep Anomaly Detection

Model ReleasesDGX agent

arXiv:2604.09166v1 Announce Type: new Abstract: Anomaly detection (AD) in chemical processes based on deep learning offers significant opportunities but requires large, diverse, and well-annotated tra

Bharat Scene Text: A Novel Comprehensive Dataset and Benchmark for Indian Language Scene Text Understanding

Model ReleasesDGX agent

arXiv:2511.23071v2 Announce Type: replace-cross Abstract: Reading scene text, that is, text appearing in images, has numerous application areas, including assistive technology, search, and e-commerce.

Claude code skill for neurotech/BCI machine learning [P]

Model ReleasesDGX agent

This r/MachineLearning post discusses a Claude Code skill tailored for neurotechnology and brain-computer interface (BCI) machine learning workflows, likely covering domain-specific tasks such as neur

Cross-Lingual Attention Distillation with Personality-Informed Generative Augmentation for Multilingual Personality Recognition

Model ReleasesDGX agent

arXiv:2604.08851v1 Announce Type: new Abstract: While significant work has been done on personality recognition, the lack of multilingual datasets remains an unresolved challenge. To address this, we

Detection and Characterization of Coordinated Online Behavior: A Survey

TutorialsDGX agent

arXiv:2408.01257v2 Announce Type: replace-cross Abstract: Coordination is a fundamental aspect of life. The advent of social media has made it integral also to online human interactions, such as those

Do you treat ChatGPT like a friend or just a tool?

IndustryDGX agent

A Reddit thread from r/ChatGPT in which users discuss their personal relationship with ChatGPT — whether they interact with it in a casual, conversational, and emotionally engaged way (as they might w

Does anyone else find themselves anthromorphizing ChatGPT?

IndustryDGX agent

This Reddit thread from r/ChatGPT invites community discussion around the common experience of anthropomorphizing ChatGPT — that is, attributing human-like qualities, emotions, or intentions to the AI

Dream to Fly: Model-Based Reinforcement Learning for Vision-Based Drone Flight

Model ReleasesDGX agent

arXiv:2501.14377v2 Announce Type: replace Abstract: Autonomous drone racing has risen as a challenging robotic benchmark for testing the limits of learning, perception, planning, and control. Expert h

DSVTLA: Deep Swin Vision Transformer-Based Transfer Learning Architecture for Multi-Type Cancer Histopathological Cancer Image Classification

Model ReleasesDGX agent

arXiv:2604.09468v1 Announce Type: cross Abstract: In this study, we proposed a deep Swin-Vision Transformer-based transfer learning architecture for robust multi-cancer histopathological image classif

Efficient Spatial-Temporal Focal Adapter with SSM for Temporal Action Detection

ApplicationsDGX agent

arXiv:2604.09164v1 Announce Type: new Abstract: Temporal human action detection aims to identify and localize action segments within untrimmed videos, serving as a pivotal task in video understanding.

Extrapolating Volition with Recursive Information Markets

SafetyDGX agent

arXiv:2604.08606v1 Announce Type: cross Abstract: One of the impediments to the efficiency of information markets is the inherent information asymmetry present in them, exacerbated by the 'buyer's ins

Finite-Sample Analysis of Nonlinear Independent Component Analysis:Sample Complexity and Identifiability Bounds

Model ReleasesDGX agent

arXiv:2604.08850v1 Announce Type: new Abstract: Independent Component Analysis (ICA) is a fundamental unsupervised learning technique foruncovering latent structure in data by separating mixed signals

FIT-GNN: Faster Inference Time for GNNs that 'FIT' in Memory Using Coarsening

Model ReleasesDGX agent

arXiv:2410.15001v5 Announce Type: replace Abstract: Scalability of Graph Neural Networks (GNNs) remains a significant challenge. To tackle this, methods like coarsening, condensation, and computation

From Paper to Program: Accelerating Quantum Many-Body Algorithm Development via a Multi-Stage LLM-Assisted Workflow

Model ReleasesDGX agent

arXiv:2604.04089v2 Announce Type: replace-cross Abstract: Large language models (LLMs) can generate code rapidly but remain unreliable for scientific algorithms whose correctness depends on structural

Get docs by MCP vs Web search

AgentsDGX agent

This Reddit post from r/ChatGPT discusses the trade-offs between retrieving documentation via an MCP (Model Context Protocol) server versus using traditional web search within AI workflows. The core c

'how can i help?' i analyzed 100 asks from portfolio companies identified in notes and emails to see what they asked 26% specific person int…

AgentsDGX agent

'how can i help?' i analyzed 100 asks from portfolio companies identified in notes and emails to see what they asked 26% specific person intro 9% investor discovery 7% hiring/talent 9% pr/media 11% bu

Identification and Anonymization of Named Entities in Unstructured Information Sources for Use in Social Engineering Detection

ApplicationsDGX agent

arXiv:2604.09016v1 Announce Type: cross Abstract: This study addresses the challenge of creating datasets for cybercrime analysis while complying with the requirements of regulations such as the Gener

Is it possible to rank in both SERP and AI Overviews?

IndustryDGX agent

Yes, it is possible to rank in both traditional SERP results and Google AI Overviews simultaneously. Google pulls content sources from its index before determining rankings or SERP features, meaning t

Law firms say lawyers are spending more time responding to swaths of AI-generated client documents, potentially leading firms to raise fixed-fee contract prices (Elizabeth Bratton/Financial Times)

IndustryDGX agent

Elizabeth Bratton / Financial Times: Law firms say lawyers are spending more time responding to swaths of AI-generated client documents, potentially leading firms to raise fixed-fee contract prices —

Mamba-Based Graph Convolutional Networks: Tackling Over-smoothing with Selective State Space

Model ReleasesDGX agent

arXiv:2501.15461v4 Announce Type: replace Abstract: Graph Neural Networks (GNNs) have shown great success in various graph-based learning tasks. However, it often faces the issue of over-smoothing as

MARINER: A 3E-Driven Benchmark for Fine-Grained Perception and Complex Reasoning in Open-Water Environments

Model ReleasesDGX agent

arXiv:2604.08615v1 Announce Type: cross Abstract: Fine-grained visual understanding and high-level reasoning in real-world open-water environments remain under-explored due to the lack of dedicated be

Maybe hot take - I’ve read a bunch of RL for image generation papers over last few months and honestly it’s been pretty disappointing. All o…

TutorialsDGX agent

Maybe hot take - I’ve read a bunch of RL for image generation papers over last few months and honestly it’s been pretty disappointing. All of them are variations of GRPO and all of them are incrementa

Mitigating Extrinsic Gender Bias for Bangla Classification Tasks

Model ReleasesDGX agent

arXiv:2411.10636v2 Announce Type: replace-cross Abstract: In this study, we investigate extrinsic gender bias in Bangla pretrained language models, a largely underexplored area in low-resource languag

NyayaMind- A Framework for Transparent Legal Reasoning and Judgment Prediction in the Indian Legal System

SafetyDGX agent

arXiv:2604.09069v1 Announce Type: cross Abstract: Court Judgment Prediction and Explanation (CJPE) aims to predict a judicial decision and provide a legally grounded explanation for a given case based

Oracle expands its partnership with fuel cell maker Bloom to procure up to 2.8 GW of capacity, after receiving a warrant to purchase $400M of Bloom stock (Jordan Novet/CNBC)

IndustryDGX agent

Jordan Novet / CNBC: Oracle expands its partnership with fuel cell maker Bloom to procure up to 2.8 GW of capacity, after receiving a warrant to purchase $400M of Bloom stock — Oracle is poised to mak

Revitalizing Black-Box Interpretability: Actionable Interpretability for LLMs via Proxy Models

Local AiDGX agent

arXiv:2505.12509v3 Announce Type: replace-cross Abstract: Post-hoc explanations provide transparency and are essential for guiding model optimization, such as prompt engineering and data sanitation. H

Scaling model size is hitting diminishing returns. The real gains are in orchestration. Our Co-Founder & Co-CEO @yshoham makes the case in a…

Model ReleasesDGX agent

Scaling model size is hitting diminishing returns. The real gains are in orchestration. Our Co-Founder & Co-CEO @yshoham makes the case in a rare long-form profile by @Calcalistech today. The man tryi

Sentiment Classification of Gaza War Headlines: A Comparative Analysis of Large Language Models and Arabic Fine-Tuned BERT Models

Model ReleasesDGX agent

arXiv:2604.08566v1 Announce Type: new Abstract: This study examines how different artificial intelligence architectures interpret sentiment in conflict-related media discourse, using the 2023 Gaza War

Shenzhen-listed server PCB maker Victory Giant plans an April 21 Hong Kong listing, aiming to raise as much as 2.2B; the company was valued at 37B on April 10 (Bloomberg)

IndustryDGX agent

Bloomberg: Shenzhen-listed server PCB maker Victory Giant plans an April 21 Hong Kong listing, aiming to raise as much as 2.2B; the company was valued at 37B on April 10 — Victory Giant Technology Hui

So the concern over Mythos and cybersecurity seems warranted.

Model ReleasesDGX agent

So the concern over Mythos and cybersecurity seems warranted. We conducted cyber evaluations of Claude Mythos Preview and found that it is the first model to complete an AISI cyber range end-to-end. 🧵

Sources: SoftBank, Sony, Honda, and six other Japanese companies launch a new AI company to develop a 1T-parameter foundation model for 'physical AI' by 2030 (Natsuki Yamamoto/Nikkei Asia)

Model ReleasesDGX agent

Natsuki Yamamoto / Nikkei Asia: Sources: SoftBank, Sony, Honda, and six other Japanese companies launch a new AI company to develop a 1T-parameter foundation model for “physical AI” by 2030 — TOKYO —

SynDocDis: A Metadata-Driven Framework for Generating Synthetic Physician Discussions Using Large Language Models

ApplicationsDGX agent

arXiv:2604.08555v1 Announce Type: new Abstract: Physician-physician discussions of patient cases represent a rich source of clinical knowledge and reasoning that could feed AI agents to enrich and eve

← Previous
1…423424425426427428
Next →