AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,737 results
10 Apr 2026

Chat SDK adds Zernio support

ToolsDGX agent

Vercel's Chat SDK has added support for Zernio, a unified social media API, via a new official adapter built and maintained by the Zernio team. Using this adapter, teams can build bots that operate...

Chat SDK now supports concurrent message handling

ToolsDGX agent

The Chat SDK has been updated to support concurrent message handling. This feature allows developers to control the processing of newly arrived messages relative to those already being handled, ens...

Chat SDK now supports scheduled Slack messages

ToolsDGX agent

The Chat SDK now enables developers to schedule Slack messages for future delivery. This functionality utilizes `thread.schedule()` by passing both the desired message content and a specific `postA...

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Chatbot-Based Assessment of Code Understanding in Automated Programming Assessment Systems

Local AiDGX agent

arXiv:2604.07304v1 Announce Type: cross Abstract: Large Language Models (LLMs) challenge conventional automated programming assessment because students can now produce functionally correct code withou

CodecSight: Leveraging Video Codec Signals for Efficient Streaming VLM Inference

HardwareDGX agent

arXiv:2604.06036v3 Announce Type: replace-cross Abstract: Video streaming analytics is a crucial workload for vision-language model serving, but the high cost of multimodal inference limits scalabilit

ConvoLearn: A Dataset for Fine-Tuning Dialogic AI Tutors

Model ReleasesDGX agent

arXiv:2601.08950v3 Announce Type: replace Abstract: Despite their growing adoption in education, LLMs remain misaligned with the core principle of effective tutoring: the dialogic construction of know

Data Leakage in Automotive Perception: Practitioners' Insights

SafetyDGX agent

arXiv:2604.06899v1 Announce Type: cross Abstract: Data leakage is the inadvertent transfer of information between training and evaluation datasets that poses a subtle, yet critical, risk to the reliab

Decompose, Look, and Reason: Reinforced Latent Reasoning for VLMs

SafetyDGX agent

arXiv:2604.07518v1 Announce Type: new Abstract: Vision-Language Models often struggle with complex visual reasoning due to the visual information loss in textual CoT. Existing methods either add the c

Demystifying OPD: Length Inflation and Stabilization Strategies for Large Language Models

SafetyDGX agent

arXiv:2604.08527v1 Announce Type: new Abstract: On-policy distillation (OPD) trains student models under their own induced distribution while leveraging supervision from stronger teachers. We identify

EgoVerse: An Egocentric Human Dataset for Robot Learning from Around the World

SafetyDGX agent

arXiv:2604.07607v1 Announce Type: cross Abstract: Robot learning increasingly depends on large and diverse data, yet robot data collection remains expensive and difficult to scale. Egocentric human da

Elastic build machines now available in beta

ToolsDGX agent

Vercel has launched Elastic build machines in beta for all paid plans, giving teams control over build performance without project-level micromanagement. Elastic build machines auto-scale based on...

Foundry: Template-Based CUDA Graph Context Materialization for Fast LLM Serving Cold Start

HardwareDGX agent

arXiv:2604.06664v1 Announce Type: cross Abstract: Modern LLM service providers increasingly rely on autoscaling and parallelism reconfiguration to respond to rapidly changing workloads, but cold-start

FP4 Explore, BF16 Train: Diffusion Reinforcement Learning via Efficient Rollout Scaling

SafetyDGX agent

arXiv:2604.06916v1 Announce Type: cross Abstract: Reinforcement-Learning-based post-training has recently emerged as a promising paradigm for aligning text-to-image diffusion models with human prefere

How Psychological Learning Paradigms Shaped and Constrained Artificial Intelligence

SafetyDGX agent

arXiv:2603.18203v3 Announce Type: replace Abstract: Current artificial intelligence systems struggle with systematic compositional reasoning: the capacity to recombine known components in novel config

https://docs.pinecone.io/integrations/gemini-cli

Model ReleasesDGX agent

I was unable to retrieve the specific content from the Pinecone Gemini CLI documentation page (`https://docs.pinecone.io/integrations/gemini-cli`) or the linked X (Twitter) post, as neither was ret...

HumanLLM: Benchmarking and Improving LLM Anthropomorphism via Human Cognitive Patterns

SafetyDGX agent

arXiv:2601.10198v3 Announce Type: replace Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities in reasoning and generation, serving as the foundation for advanced persona s

if it creates a skill, and it errors, it will (at least sometimes) just try to fix it. what’s cool is that all the web search and code writi…

Model ReleasesDGX agent

if it creates a skill, and it errors, it will (at least sometimes) just try to fix it. what’s cool is that all the web search and code writing is offloaded to Claude (or whatever chat you’re using), a

In a world where writing code to build websites and apps is trivial (thank you Lovable, Cursor, Claude,...), the real differentiation for yo…

Model ReleasesDGX agent

In a world where writing code to build websites and apps is trivial (thank you Lovable, Cursor, Claude,...), the real differentiation for you and your company (and what makes you successful) will be h

k-server-bench: Automating Potential Discovery for the k-Server Conjecture

Model ReleasesDGX agent

arXiv:2604.07240v1 Announce Type: cross Abstract: We introduce a code-based challenge for automated, open-ended mathematical discovery based on the k-server conjecture, a central open problem in com

Learning Who Disagrees: Demographic Importance Weighting for Modeling Annotator Distributions with DiADEM

SafetyDGX agent

arXiv:2604.08425v1 Announce Type: cross Abstract: When humans label subjective content, they disagree, and that disagreement is not noise. It reflects genuine differences in perspective shaped by anno

LINE: LLM-based Iterative Neuron Explanations for Vision Models

SafetyDGX agent

arXiv:2604.08039v1 Announce Type: new Abstract: Interpreting the concepts encoded by individual neurons in deep neural networks is a crucial step towards understanding their complex decision-making pr

LiteParse hit 4K+ GitHub stars in 3 weeks. ~500 pages in 2 seconds. No GPU. No API keys. 50+ file formats. Now @LoganMarkewich, our Head of …

HardwareDGX agent

LiteParse hit 4K+ GitHub stars in 3 weeks. ~500 pages in 2 seconds. No GPU. No API keys. 50+ file formats. Now @LoganMarkewich, our Head of Open Source, will show you how to build with it. Live worksh

LLM-based Schema-Guided Extraction and Validation of Missing-Person Intelligence from Heterogeneous Data Sources

SafetyDGX agent

arXiv:2604.06571v1 Announce Type: cross Abstract: Missing-person and child-safety investigations rely on heterogeneous case documents, including structured forms, bulletin-style posters, and narrative

Mark “Metaverse” Zuckerberg totally bought Moltbook at the peak of the market 🤣

SafetyDGX agent

Meta acquired Moltbook, an AI-only social network launched in January 2026 by entrepreneurs Matt Schlicht and Ben Parr, after the platform had gone viral and then faded in popularity. Moltbook is ...

MDP modeling for multi-stage stochastic programs

SafetyDGX agent

arXiv:2509.22981v2 Announce Type: replace Abstract: We study a class of multi-stage stochastic programs, which incorporate modeling features from Markov decision processes (MDPs). This class includes

MedDialBench: Benchmarking LLM Diagnostic Robustness under Parametric Adversarial Patient Behaviors

Model ReleasesDGX agent

arXiv:2604.06846v1 Announce Type: cross Abstract: Interactive medical dialogue benchmarks have shown that LLM diagnostic accuracy degrades significantly when interacting with non-cooperative patients,

MonoUNet: A Robust Tiny Neural Network for Automated Knee Cartilage Segmentation on Point-of-Care Ultrasound Devices

SafetyDGX agent

arXiv:2604.07780v1 Announce Type: cross Abstract: Objective: To develop a robust and compact deep learning model for automated knee cartilage segmentation on point-of-care ultrasound (POCUS) devices.

Multi-Faceted Self-Consistent Preference Alignment for Query Rewriting in Conversational Search

SafetyDGX agent

arXiv:2604.06771v1 Announce Type: cross Abstract: Conversational Query Rewriting (CQR) aims to rewrite ambiguous queries to achieve more efficient conversational search. Early studies have predominant

Not All Tokens See Equally: Perception-Grounded Policy Optimization for Large Vision-Language Models

Model ReleasesDGX agent

arXiv:2604.01840v2 Announce Type: replace Abstract: While Reinforcement Learning from Verifiable Rewards (RLVR) has advanced reasoning in Large Vision-Language Models (LVLMs), prevailing frameworks su

OpenVLThinkerV2: A Generalist Multimodal Reasoning Model for Multi-domain Visual Tasks

SafetyDGX agent

arXiv:2604.08539v1 Announce Type: cross Abstract: Group Relative Policy Optimization (GRPO) has emerged as the de facto Reinforcement Learning (RL) objective driving recent advancements in Multimodal

People ask me how I choose what model to route queries to. It's simple. Claude gets knowledge work anything below that would be degrading Cl…

Model ReleasesDGX agent

I was unable to retrieve the specific tweet content from that URL, as the post requires a logged-in X (Twitter) session to access, and search results did not surface the full text of that specific ...

PeReGrINE: Evaluating Personalized Review Fidelity with User Item Graph Context

Model ReleasesDGX agent

arXiv:2604.07788v1 Announce Type: cross Abstract: We introduce PeReGrINE, a benchmark and evaluation framework for personalized review generation grounded in graph-structured user--item evidence. PeRe

Phantasia: Context-Adaptive Backdoors in Vision Language Models

ResearchDGX agent

arXiv:2604.08395v1 Announce Type: new Abstract: Recent advances in Vision-Language Models (VLMs) have greatly enhanced the integration of visual perception and linguistic reasoning, driving rapid prog

PIArena: A Platform for Prompt Injection Evaluation

ApplicationsDGX agent

arXiv:2604.08499v1 Announce Type: cross Abstract: Prompt injection attacks pose serious security risks across a wide range of real-world applications. While receiving increasing attention, the communi

Plug-and-Play Logit Fusion for Heterogeneous Pathology Foundation Models

SafetyDGX agent

arXiv:2604.07779v1 Announce Type: new Abstract: Pathology foundation models (FMs) have become central to computational histopathology, offering strong transfer performance across a wide range of diagn

🚀 Qwen Code v0.14.0 – v0.14.2 are now available Channels:Control Qwen Code remotely from Telegram, DingTalk, or WeChat — send a message fro…

Model ReleasesDGX agent

🚀 Qwen Code v0.14.0 – v0.14.2 are now available Channels:Control Qwen Code remotely from Telegram, DingTalk, or WeChat — send a message from your phone, get results on your server Cron Jobs :Schedule

Raising the security baseline: Essential AI and cloud security now on by default

Model ReleasesDGX agent

The rapid evolution of AI is redefining industries, while also exposing organizations to new risks. At Google Cloud, we believe that modern cloud defense should have AI protection built in and accessi

ReflectRM: Boosting Generative Reward Models via Self-Reflection within a Unified Judgment Framework

SafetyDGX agent

arXiv:2604.07506v1 Announce Type: cross Abstract: Reward Models (RMs) are critical components in the Reinforcement Learning from Human Feedback (RLHF) pipeline, directly determining the alignment qual

Replit now deploys directly to Databricks. Your apps run inside your Databricks environment while inheriting its security, governance, and d…

ToolsDGX agent

Replit now deploys directly to Databricks. Your apps run inside your Databricks environment while inheriting its security, governance, and data access. Beta is live. Enterprises are already building w

Reset-Free Reinforcement Learning for Real-World Agile Driving: An Empirical Study

SafetyDGX agent

arXiv:2604.07672v1 Announce Type: new Abstract: This paper presents an empirical study of reset-free reinforcement learning (RL) for real-world agile driving, in which a physical 1/10-scale vehicle le

RLBoost: Harvesting Preemptible Resources for Cost-Efficient Reinforcement Learning on LLMs

HardwareDGX agent

arXiv:2510.19225v3 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has become essential for unlocking advanced reasoning capabilities in large language models (LLMs). RL workflows i

RoSHI: A Versatile Robot-oriented Suit for Human Data In-the-Wild

SafetyDGX agent

arXiv:2604.07331v1 Announce Type: cross Abstract: Scaling up robot learning will likely require human data containing rich and long-horizon interactions in the wild. Existing approaches for collecting

SE-Enhanced ViT and BiLSTM-Based Intrusion Detection for Secure IIoT and IoMT Environments

Model ReleasesDGX agent

arXiv:2604.06254v1 Announce Type: cross Abstract: With the rapid growth of interconnected devices in Industrial and Medical Internet of Things (IIoT and MIoT) ecosystems, ensuring timely and accurate

Self-Debias: Self-correcting for Debiasing Large Language Models

SafetyDGX agent

arXiv:2604.08243v1 Announce Type: new Abstract: Although Large Language Models (LLMs) demonstrate remarkable reasoning capabilities, inherent social biases often cascade throughout the Chain-of-Though

SeMoBridge: Semantic Modality Bridge for Efficient Few-Shot Adaptation of CLIP

SafetyDGX agent

arXiv:2509.26036v3 Announce Type: replace Abstract: While Contrastive Language-Image Pretraining (CLIP) excels at zero-shot tasks by aligning image and text embeddings, its performance in few-shot cla

SIM1: Physics-Aligned Simulator as Zero-Shot Data Scaler in Deformable Worlds

SafetyDGX agent

arXiv:2604.08544v1 Announce Type: cross Abstract: Robotic manipulation with deformable objects represents a data-intensive regime in embodied learning, where shape, contact, and topology co-evolve in

Steering the Verifiability of Multimodal AI Hallucinations

TutorialsDGX agent

arXiv:2604.06714v1 Announce Type: new Abstract: AI applications driven by multimodal large language models (MLLMs) are prone to hallucinations and pose considerable risks to human users. Crucially, su

Stop Listening to Me! How Multi-turn Conversations Can Degrade LLM Diagnostic Reasoning

ApplicationsDGX agent

arXiv:2603.11394v2 Announce Type: replace Abstract: Patients and clinicians are increasingly using chatbots powered by large language models (LLMs) for healthcare inquiries. While state-of-the-art LLM

Summary of CVE-2026-23869

ToolsDGX agent

CVE-2026-23869 is a high-severity Denial of Service vulnerability affecting React Server Components. It can lead to Denial of Service and is present in Next.js 13.x, 14.x, 15.x, and 16.x, impactin...

SymptomWise: A Deterministic Reasoning Layer for Reliable and Efficient AI Systems

SafetyDGX agent

arXiv:2604.06375v1 Announce Type: new Abstract: AI-driven symptom analysis systems face persistent challenges in reliability, interpretability, and hallucination. End-to-end generative approaches ofte

SYN-DIGITS: A Synthetic Control Framework for Calibrated Digital Twin Simulation

SafetyDGX agent

arXiv:2604.07513v1 Announce Type: cross Abstract: AI-based persona simulation -- often referred to as digital twin simulation -- is increasingly used for market research, recommender systems, and soci

The AI Skills Shift: Mapping Skill Obsolescence, Emergence, and Transition Pathways in the LLM Era

Model ReleasesDGX agent

arXiv:2604.06906v1 Announce Type: cross Abstract: As Large Language Models reshape the global labor market, policymakers and workers need empirical data on which occupational skills may be most suscep

The Art of (Mis)alignment: How Fine-Tuning Methods Effectively Misalign and Realign LLMs in Post-Training

SafetyDGX agent

arXiv:2604.07754v1 Announce Type: cross Abstract: The deployment of large language models (LLMs) raises significant ethical and safety concerns. While LLM alignment techniques are adopted to improve m

The Sustainability Gap in Robotics: A Large-Scale Survey of Sustainability Awareness in 50,000 Research Articles

SafetyDGX agent

arXiv:2604.07921v1 Announce Type: new Abstract: We present a large-scale survey of sustainability communication and motivation in robotics research. Our analysis covers nearly 50,000 open-access paper

Towards the Development of an LLM-Based Methodology for Automated Security Profiling in Compliance with Ukrainian Cybersecurity Regulations

SafetyDGX agent

arXiv:2604.06274v1 Announce Type: cross Abstract: In recent years, the pace of development of information technology in various areas has increased drastically, forcing cybersecurity specialists to co

TSUBASA: Improving Long-Horizon Personalization via Evolving Memory and Self-Learning with Context Distillation

Model ReleasesDGX agent

arXiv:2604.07894v1 Announce Type: new Abstract: Personalized large language models (PLLMs) have garnered significant attention for their ability to align outputs with individual's needs and preference

Vercel plugin now supported on OpenAI Codex and Codex CLI

ToolsDGX agent

The Vercel plugin now supports OpenAI Codex and Codex CLI, expanding on its initial launch with Claude Code and Cursor. With the plugin, teams can access over 39 platform skills, three specialist ...

WASD: Locating Critical Neurons as Sufficient Conditions for Explaining and Controlling LLM Behavior

Model ReleasesDGX agent

arXiv:2603.18474v2 Announce Type: replace Abstract: Precise behavioral control of large language models (LLMs) is critical for complex applications. However, existing methods often incur high training

What Drives Representation Steering? A Mechanistic Case Study on Steering Refusal

SafetyDGX agent

arXiv:2604.08524v1 Announce Type: cross Abstract: Applying steering vectors to large language models (LLMs) is an efficient and effective model alignment technique, but we lack an interpretable explan

Zero-configuration Django support

ToolsDGX agent

Vercel now supports Django with zero configuration, enabling developers to instantly deploy Django full-stack applications or APIs without any manual setup. Vercel automatically detects Django's `m...

← Previous
1…258259260261262…296
Next →