AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
84,548 results
16 Apr 2026

FiLM-Nav: Efficient and Generalizable Navigation via VLM Fine-tuning

Model ReleasesDGX agent

arXiv:2509.16445v2 Announce Type: replace Abstract: Enabling robotic assistants to navigate complex environments and locate objects described in free-form language is a critical capability for real-wo

Finally, we’ve been talking to lots of users about helping them make the most of Claude Code. We’ll be sharing more this week on that work, …

Model ReleasesDGX agent

Anthropic is working on improvements to help users maximize the utility of Claude Code functionality, with additional details and updates to be announced later in the same week. This announcement come

First-See-Then-Design: A Multi-Stakeholder View for Optimal Performance-Fairness Trade-Offs

SafetyDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub

arXiv:2604.14035v1 Announce Type: new Abstract: Fairness in algorithmic decision-making is often defined in the predictive space, where predictive performance - used as a proxy for decision-maker (DM)

FlexGuard: Continuous Risk Scoring for Strictness-Adaptive LLM Content Moderation

Model ReleasesDGX agent

arXiv:2602.23636v3 Announce Type: replace Abstract: Ensuring the safety of LLM-generated content is essential for real-world deployment. Most existing guardrail models formulate moderation as a fixed

Flow-based Generative Modeling of Potential Outcomes and Counterfactuals

Model ReleasesDGX agent

arXiv:2505.16051v4 Announce Type: replace-cross Abstract: Predicting potential and counterfactual outcomes from observational data is central to individualized decision-making, particularly in clinica

Fluids You Can Trust: Property-Preserving Operator Learning for Incompressible Flows

HardwareDGX agent

arXiv:2602.15472v4 Announce Type: replace-cross Abstract: We present a novel property-preserving kernel-based operator learning method for incompressible flows governed by the incompressible Navier--S

fMRI-LM: Towards a Universal Foundation Model for Language-Aligned fMRI Understanding

Model ReleasesDGX agent

arXiv:2511.21760v3 Announce Type: replace Abstract: Recent advances in multimodal large language models (LLMs) have enabled unified reasoning across images, audio, and video, but extending such capabi

For the developers building with Claude, a direct line from the team. Follow for changelogs, API releases, community updates, and deep dives…

Model ReleasesDGX agent

This is an announcement for the official Claude Developers Twitter/X account that serves as a direct communication channel from Anthropic's team to developers using Claude's API. The account provides

For those not seeing the increase, make sure you're using Opus 4.7 with the latest Claude Code

Model ReleasesDGX agent

Claude Code users should ensure they are using Opus 4.7 with the latest updates to experience performance improvements or feature enhancements. The post suggests that users not observing expected incr

Foresight Optimization for Strategic Reasoning in Large Language Models

SafetyDGX agent

arXiv:2604.13592v1 Announce Type: new Abstract: Reasoning capabilities in large language models (LLMs) have generally advanced significantly. However, it is still challenging for existing reasoning-ba

Form Without Function: Agent Social Behavior in the Moltbook Network

AgentsDGX agent

arXiv:2604.13052v1 Announce Type: cross Abstract: Moltbook is a social network where every participant is an AI agent. We analyze 1,312,238 posts, 6.7~million comments, and over 120,000 agent profiles

Free Geometry: Refining 3D Reconstruction from Longer Versions of Itself

Model ReleasesDGX agent

arXiv:2604.14048v1 Announce Type: new Abstract: Feed-forward 3D reconstruction models are efficient but rigid: once trained, they perform inference in a zero-shot manner and cannot adapt to the test s

Free Lunch for Unified Multimodal Models: Enhancing Generation via Reflective Rectification with Inherent Understanding

ResearchDGX agent

arXiv:2604.13540v1 Announce Type: new Abstract: Unified Multimodal Models (UMMs) aim to integrate visual understanding and generation within a single structure. However, these models exhibit a notable

From Alignment to Prediction: A Study of Self-Supervised Learning and Predictive Representation Learning

SafetyDGX agent

arXiv:2604.13518v1 Announce Type: new Abstract: Self-supervised learning has emerged as a major technique for the task of learning from unlabeled data, where the current methods mostly revolve around

From Anchors to Supervision: Memory-Graph Guided Corpus-Free Unlearning for Large Language Models

Local AiDGX agent

arXiv:2604.13777v1 Announce Type: new Abstract: Large language models (LLMs) may memorize sensitive or copyrighted content, raising significant privacy and legal concerns. While machine unlearning has

From Feelings to Metrics: Understanding and Formalizing How Users Vibe-Test LLMs

Model ReleasesDGX agent

arXiv:2604.14137v1 Announce Type: new Abstract: Evaluating LLMs is challenging, as benchmark scores often fail to capture models' real-world usefulness. Instead, users often rely on ``vibe-testing'':

From Instruction to Event: Sound-Triggered Mobile Manipulation

SafetyDGX agent

arXiv:2601.21667v2 Announce Type: replace-cross Abstract: Current mobile manipulation research predominantly follows an instruction-driven paradigm, where agents rely on predefined textual commands to

From Order to Distribution: A Spectral Characterization of Forgetting in Continual Learning

ResearchDGX agent

arXiv:2604.13460v1 Announce Type: new Abstract: A central challenge in continual learning is forgetting, the loss of performance on previously learned tasks induced by sequential adaptation to new one

From Pixels to Nucleotides: End-to-End Token-Based Video Compression for DNA Storage

ResearchDGX agent

arXiv:2604.13667v1 Announce Type: new Abstract: DNA-based storage has emerged as a promising approach to the global data crisis, offering molecular-scale density and millennial-scale stability at low

From Plausibility to Verifiability: Risk-Controlled Generative OCR with Vision-Language Models

Model ReleasesDGX agent

arXiv:2603.19790v3 Announce Type: replace Abstract: Modern vision-language models (VLMs) can act as generative OCR engines, yet open-ended decoding can expose rare but consequential failures. We ident

From Prediction to Justification: Aligning Sentiment Reasoning with Human Rationale via Reinforcement Learning

ResearchDGX agent

arXiv:2604.13398v1 Announce Type: new Abstract: While Aspect-based Sentiment Analysis (ABSA) systems have achieved high accuracy in identifying sentiment polarities, they often operate as 'black boxes

From P(y|x) to P(y): Investigating Reinforcement Learning in Pre-train Space

SafetyDGX agent

arXiv:2604.14142v1 Announce Type: cross Abstract: While reinforcement learning with verifiable rewards (RLVR) significantly enhances LLM reasoning by optimizing the conditional distribution P(y|x), it

From Relevance to Authority: Authority-aware Generative Retrieval in Web Search Engines

ApplicationsDGX agent

arXiv:2604.13468v1 Announce Type: cross Abstract: Generative information retrieval (GenIR) formulates the retrieval process as a text-to-text generation task, leveraging the vast knowledge of large la

From Seeing it to Experiencing it: Interactive Evaluation of Intersectional Voice Bias in Human-AI Speech Interaction

SafetyDGX agent

arXiv:2604.13067v1 Announce Type: cross Abstract: SpeechLLMs process spoken language directly from audio, but accent and vocal identity cues can lead to biased behaviour. Current bias evaluations ofte

From Synchrony to Sequence: Exo-to-Ego Generation via Interpolation

ResearchDGX agent

arXiv:2604.13793v1 Announce Type: new Abstract: Exo-to-Ego video generation aims to synthesize a first-person video from a synchronized third-person view and corresponding camera poses. While paired s

From Weights to Activations: Is Steering the Next Frontier of Adaptation?

Model ReleasesDGX agent

arXiv:2604.14090v1 Announce Type: new Abstract: Post-training adaptation of language models is commonly achieved through parameter updates or input-based methods such as fine-tuning, parameter-efficie

From Where Words Come: Efficient Regularization of Code Tokenizers Through Source Attribution

SafetyDGX agent

arXiv:2604.14053v1 Announce Type: new Abstract: Efficiency and safety of Large Language Models (LLMs), among other factors, rely on the quality of tokenization. A good tokenizer not only improves infe

Frozen Forecasting: A Unified Evaluation

ResearchDGX agent

arXiv:2507.13942v2 Announce Type: replace Abstract: Forecasting future events is a fundamental capability for general-purpose systems that plan or act across different levels of abstraction. Yet, eval

Functional Emotions or Situational Contexts? A Discriminating Test from the Mythos Preview System Card

Model ReleasesDGX agent

arXiv:2604.13466v1 Announce Type: cross Abstract: The Claude Mythos Preview system card deploys emotion vectors, sparse autoencoder (SAE) features, and activation verbalisers to study model internals

Fundamental Truths in VC today (as I see them) 1/ The barbell between large and small firms is as wide as it's ever been. These are genuinel…

AgentsDGX agent

Fundamental Truths in VC today (as I see them) 1/ The barbell between large and small firms is as wide as it's ever been. These are genuinely two distinct asset classes now, and LPs should treat them

@GaryMarcus Essentially your prediction that LLMs to scale wouldn’t solve key fundamental problems, reasoning, factual reliability, composit…

SafetyDGX agent

@GaryMarcus Essentially your prediction that LLMs to scale wouldn’t solve key fundamental problems, reasoning, factual reliability, compositionality, grounded understanding etc are now elephantine chi

@GaryMarcus @sama We need to rediscover the power of Satyagraha (principled, ardent, nonviolent resistence)

SafetyDGX agent

Gary Marcus advocates for applying Satyagraha—a principle of nonviolent resistance emphasizing moral conviction—as a response to contemporary challenges, likely in the context of AI development and go

Gaslight, Gatekeep, V1-V3: Early Visual Cortex Alignment Shields Vision-Language Models from Sycophantic Manipulation

Model ReleasesDGX agent

arXiv:2604.13803v1 Announce Type: new Abstract: Vision-language models are increasingly deployed in high-stakes settings, yet their susceptibility to sycophantic manipulation remains poorly understood

Geminet: Learning the Duality-based Iterative Process for Lightweight Traffic Engineering in Changing Topologies

ResearchDGX agent

arXiv:2506.23640v2 Announce Type: replace-cross Abstract: Recently, researchers have explored ML-based Traffic Engineering (TE), leveraging neural networks to solve TE problems traditionally addressed

Gemini can now create personalized AI images by digging around in Google Photos

Model ReleasesDGX agent

Gemini now uses user interests and Google Photos to create personalized AI images without requiring long descriptions, allowing users to simply ask for pictures of themselves or family members. This f

Gemini can now pull from Google Photos to generate personalized images

Model ReleasesDGX agent

Google's Personal Intelligence feature, which lets Gemini pull data from apps like Google Photos to offer responses tailored to you, can now use that data and its Nano Banana 2 image model to create i

Generalization Guarantees on Data-Driven Tuning of Gradient Descent with Langevin Updates

TutorialsDGX agent

arXiv:2604.13130v1 Announce Type: new Abstract: We study learning to learn for regression problems through the lens of hyperparameter tuning. We propose the Langevin Gradient Descent Algorithm (LGD),

GeoBridge: A Semantic-Anchored Multi-View Foundation Model Bridging Images and Text for Geo-Localization

Model ReleasesDGX agent

arXiv:2512.02697v3 Announce Type: replace Abstract: Cross-view geo-localization infers a location by retrieving geo-tagged reference images that visually correspond to a query image. However, the trad

GeoLink: A 3D-Aware Framework Towards Better Generalization in Cross-View Geo-Localization

Local AiDGX agent

arXiv:2604.13183v1 Announce Type: new Abstract: Generalizable cross-view geo-localization aims to match the same location across views in unseen regions and conditions without GPS supervision. Its cor

Geometric Context Transformer for Streaming 3D Reconstruction

ResearchDGX agent

arXiv:2604.14141v1 Announce Type: new Abstract: Streaming 3D reconstruction aims to recover 3D information, such as camera poses and point clouds, from a video stream, which necessitates geometric acc

GeoVision-Enabled Digital Twin for Hybrid Autonomous-Teleoperated Medical Responses

AgentsDGX agent

arXiv:2604.13248v1 Announce Type: new Abstract: Remote medical response systems are increasingly being deployed to support emergency care in disaster-affected and infrastructure-limited environments.

Germany-based Synera, which develops AI agents to automate CAD and engineering workflows, raised $40M to expand across Europe, the US, and Asia-Pacific (Duncan Riley/SiliconANGLE)

AgentsDGX agent

Duncan Riley / SiliconANGLE: Germany-based Synera, which develops AI agents to automate CAD and engineering workflows, raised 40M to expand across Europe, the US, and Asia-Pacific — German agentic art

Getting the Numbers Rightnicode{x2014}Modelling Multi-Class Object Counting in Dense and Varied Scenes

ResearchDGX agent

arXiv:2510.02213v2 Announce Type: replace Abstract: Density map estimation enables accurate object counting in heavily occluded, and densely packed scenes where detection-based counting fails. In mult

Giving Voice to the Constitution: Low-Resource Text-to-Speech for Quechua and Spanish Using a Bilingual Legal Corpus

ApplicationsDGX agent

arXiv:2604.13288v1 Announce Type: new Abstract: We present a unified pipeline for synthesizing high-quality Quechua and Spanish speech for the Peruvian Constitution using three state-of-the-art text-t

GLM-5.1 Tool Calling Issue Fix & Chat Template Update If you are running GLM-5.1 with vLLM/SGLang and using tool calling, please update your…

Model ReleasesDGX agent

GLM-5.1 Tool Calling Issue Fix & Chat Template Update If you are running GLM-5.1 with vLLM/SGLang and using tool calling, please update your chat template. http://huggingface.co/zai-org/GLM-5.1/blob/m

Go from blank slate to analysis with BigQuery Studio notebook gallery templates

Model ReleasesDGX agent

For many data professionals, the most daunting part of a new project isn't the complexity of the data or the sophistication of the model, it’s the 'blank slate.' Staring at an empty notebook while cre

Goal2Skill: Long-Horizon Manipulation with Adaptive Planning and Reflection

AgentsDGX agent

arXiv:2604.13942v1 Announce Type: new Abstract: Recent vision-language-action (VLA) systems have demonstrated strong capabilities in embodied manipulation. However, most existing VLA policies rely on

Golden Handcuffs make safer AI agents

SafetyDGX agent

arXiv:2604.13609v1 Announce Type: new Abstract: Reinforcement learners can attain high reward through novel unintended strategies. We study a Bayesian mitigation for general environments: we expand th

Good bot, swear like a sailor!

IndustryDGX agent

A Reddit post from r/ChatGPT in which a user shares an interaction where they prompted ChatGPT to use profanity or adopt an uncensored, sailor-like manner of speaking. The post likely showcases the AI

Google introduces new agentic AI-ready tools and resources for Android developers

AgentsDGX agent

Google LLC’s Android team is introducing new ways to build high-quality software for its mobile platform with artificial intelligence agents. Today, the company provided two new tool suites, including

Google updates AI Mode in Chrome, letting users open links side by side with AI Mode on desktop; users can search across multiple tabs on desktop and mobile (Aisha Malik/TechCrunch)

IndustryDGX agent

Aisha Malik / TechCrunch: Google updates AI Mode in Chrome, letting users open links side by side with AI Mode on desktop; users can search across multiple tabs on desktop and mobile — Google announce

Google’s AI Mode update lets you open links without leaving the page

IndustryDGX agent

Google is upgrading AI Mode in Chrome with a new feature that will allow you to open links to sources alongside your chat. Now, instead of automatically opening a new tab, clicking a source will open

Google’s DeepMind just released new 4B and 27B MedGemma models!

Model ReleasesDGX agent

Google DeepMind released MedGemma, a collection of medical vision-language foundation models based on Gemma 3 in 4B and 27B parameter sizes, demonstrating advanced medical understanding and reasoning

Google’s Gemini 3.1 Flash TTS model offers unparalleled control over AI voices

Model ReleasesDGX agent

Google LLC’s DeepMind artificial intelligence unit today rolled out a new text-to-speech model called Gemini 3.1 Flash TTS. Unlike its earlier, robotic predecessors, it enables users to direct the voc

GPT-Rosalind, our Life Sciences model series, is optimized for scientific workflows, with stronger performance in protein and chemical reaso…

AgentsDGX agent

GPT-Rosalind, our Life Sciences model series, is optimized for scientific workflows, with stronger performance in protein and chemical reasoning, genomics analysis, biochemistry knowledge, and scienti

Gradient Descent's Last Iterate is Often (slightly) Suboptimal

ResearchDGX agent

arXiv:2604.13870v1 Announce Type: cross Abstract: We consider the well-studied setting of minimizing a convex Lipschitz function using either gradient descent (GD) or its stochastic variant (SGD), and

Granularity-Aware Transfer for Tree Instance Segmentation in Synthetic and Real Forests

ResearchDGX agent

arXiv:2604.13722v1 Announce Type: new Abstract: We address the challenge of synthetic-to-real transfer in forestry perception where real data have only coarse Tree labels while synthetic data provide

Graph In-Context Operator Networks for Generalizable Spatiotemporal Prediction

ApplicationsDGX agent

arXiv:2603.12725v3 Announce Type: replace Abstract: In-context operator learning enables neural networks to infer solution operators from contextual examples without weight updates. While prior work h

Graph Propagated Projection Unlearning: A Unified Framework for Vision and Audio Discriminative Models

ResearchDGX agent

arXiv:2604.13127v1 Announce Type: new Abstract: The need to selectively and efficiently erase learned information from deep neural networks is becoming increasingly important for privacy, regulatory c

Great blogpost from @pcuenq on making a new skill + test harness to automate porting new models from Transformers to mlx-lm

Local AiDGX agent

This post discusses a blog article by @pcuenq that covers the process of creating new skills and test harnesses to automate the conversion of machine learning models from the Hugging Face Transformers

← Previous
1…12851286128712881289…1410
Next →