AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,429 results
22 Apr 2026

On Temperature-Constrained Non-Deterministic Machine Translation: Potential and Evaluation

ApplicationsDGX agent

arXiv:2601.13729v2 Announce Type: replace Abstract: In recent years, the non-deterministic properties of language models have garnered considerable attention and have shown a significant influence on

Optimal Routing for Federated Learning over Dynamic Satellite Networks: Tractable or Not?

Local AiDGX agent

arXiv:2604.19399v1 Announce Type: new Abstract: Federated learning (FL) is a key paradigm for distributed model learning across decentralized data sources. Communication in each FL round typically con

PC2Model: ISPRS benchmark on 3D point cloud to model registration

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.19596v1 Announce Type: new Abstract: Point cloud registration involves aligning one point cloud with another or with a three-dimensional (3D) model, enabling the integration of multimodal d

Pixels or Positions? Benchmarking Modalities in Group Activity Recognition

Model ReleasesDGX agent

arXiv:2511.12606v3 Announce Type: replace Abstract: Group Activity Recognition (GAR) is well studied on the video modality for surveillance and indoor team sports (e.g., volleyball, basketball). Yet,

Position: LLM Watermarking Should Align Stakeholders' Incentives for Practical Adoption

ApplicationsDGX agent

arXiv:2510.18333v2 Announce Type: replace-cross Abstract: Despite progress in watermarking algorithms for large language models (LLMs), real-world deployment remains limited. We argue that this gap st

Proposing Topic Models and Evaluation Frameworks for Analyzing Associations with External Outcomes: An Application to Leadership Analysis Using Large-Scale Corporate Review Data

SafetyDGX agent

arXiv:2604.18919v1 Announce Type: new Abstract: Analyzing topics extracted from text data in relation to external outcomes is important across fields such as computational social science and organizat

Regression with Large Language Models for Materials and Molecular Property Prediction

Model ReleasesDGX agent

arXiv:2409.06080v2 Announce Type: replace-cross Abstract: We demonstrate the ability of large language models (LLMs) to perform material and molecular property regression tasks, a significant deviatio

Regulating Artificial Intimacy: From Locks and Blocks to Relational Accountability

SafetyDGX agent

arXiv:2604.18893v1 Announce Type: cross Abstract: A series of high-profile tragedies involving companion chatbots has triggered an unusually rapid regulatory response. Several jurisdictions, including

Relational AI in Education: Reciprocity, Participatory Design, and Indigenous Worldviews

TutorialsDGX agent

arXiv:2604.19099v1 Announce Type: cross Abstract: Education is not merely the transmission of information or the optimisation of individual performance; it is a fundamentally social, constructive, and

Rethinking Scale: Deployment Trade-offs of Small Language Models under Agent Paradigms

AgentsDGX agent

arXiv:2604.19299v1 Announce Type: cross Abstract: Despite the impressive capabilities of large language models, their substantial computational costs, latency, and privacy risks hinder their widesprea

RoLegalGEC: Legal Domain Grammatical Error Detection and Correction Dataset for Romanian

ApplicationsDGX agent

arXiv:2604.19593v1 Announce Type: cross Abstract: The importance of clear and correct text in legal documents cannot be understated, and, consequently, a grammatical error correction tool meant to ass

Safe Continual Reinforcement Learning in Non-stationary Environments

Model ReleasesDGX agent

arXiv:2604.19737v1 Announce Type: new Abstract: Reinforcement learning (RL) offers a compelling data-driven paradigm for synthesizing controllers for complex systems when accurate physical models are

SAHM: A Benchmark for Arabic Financial and Shari'ah-Compliant Reasoning

Model ReleasesDGX agent

arXiv:2604.19098v1 Announce Type: cross Abstract: English financial NLP has progressed rapidly through benchmarks for sentiment, document understanding, and financial question answering, while Arabic

Seeing Candidates at Scale: Multimodal LLMs for Visual Political Communication on Instagram

ApplicationsDGX agent

arXiv:2604.19489v1 Announce Type: new Abstract: This paper presents a computational case study that evaluates the capabilities of specialized machine learning models and emerging multimodal large lang

🆕 Shopify's AI-Native Engineering: 100% adoption, unlimited tokens, Tangle, Tangent, & SimGym https://www.latent.space/p/shopify @Shopify C…

ToolsDGX agent

🆕 Shopify's AI-Native Engineering: 100% adoption, unlimited tokens, Tangle, Tangent, & SimGym https://www.latent.space/p/shopify @Shopify CTO @MParakhin explains why near-universal AI adoption is chan

Small and midsize businesses jumpstart their AI transformations with Gemini Enterprise

Model ReleasesDGX agent

Small businesses are the backbone of the global economy. With 400 million SMBs worldwide and 36 million in the U.S. alone, they provide 50% of global employment. Now, with Google Cloud AI, they’re sca

Sparse Network Inference under Imperfect Detection and its Application to Ecological Networks

ApplicationsDGX agent

arXiv:2604.18820v1 Announce Type: cross Abstract: Recovering latent structure from count data has received considerable attention in network inference, particularly when one seeks both cross-group int

Tabloid reports linking 10 missing and dead scientists spur FBI probe

IndustryDGX agent

The FBI is looking for any connections among the recent deaths and disappearances of at least 10 scientists who had ties to government science projects or other sensitive information. At least 10 indi

The great AI democratization begins. Every startup just got a PhD-level ML team for free 🧵

IndustryDGX agent

This thread discusses how advances in AI tooling and accessible models have lowered barriers to entry for startups, enabling small teams to leverage capabilities previously requiring specialized ML ex

This has in fact (at various levels) happened more than 1000 times. Check out @DamienCharlotin’s tracker and my substack essays about it.

ApplicationsDGX agent

This has in fact (at various levels) happened more than 1000 times. Check out @DamienCharlotin’s tracker and my substack essays about it. And it begins Sullivan & Cromwell just admitted to a federal j

This is why I won't use proprietary hosted embedding models myself - I am more than happy to pay for a hosted solution (cheaper, faster and …

ToolsDGX agent

This is why I won't use proprietary hosted embedding models myself - I am more than happy to pay for a hosted solution (cheaper, faster and more convenient than self-hosting) but I want an open weight

this looks like a website but it’s an interactive generative video 🫨

AgentsDGX agent

this looks like a website but it’s an interactive generative video 🫨 Imagine every pixel on your screen, streamed live directly from a model. No HTML, no layout engine, no code. Just exactly what you

Time Series Augmented Generation for Financial Applications

Model ReleasesDGX agent

arXiv:2604.19633v1 Announce Type: new Abstract: Evaluating the reasoning capabilities of Large Language Models (LLMs) for complex, quantitative financial tasks is a critical and unsolved challenge. St

Towards Optimal Agentic Architectures for Offensive Security Tasks

Model ReleasesDGX agent

arXiv:2604.18718v1 Announce Type: cross Abstract: Agentic security systems increasingly audit live targets with tool-using LLMs, but prior systems fix a single coordination topology, leaving unclear w

Towards Reliable Human Evaluations in Gesture Generation: Insights from a Community-Driven State-of-the-Art Benchmark

Model ReleasesDGX agent

arXiv:2511.01233v3 Announce Type: replace Abstract: We review human evaluation practices in automatic, speech-driven 3D gesture generation and find a lack of standardisation and frequent use of flawed

Tstars-Tryon 1.0: Robust and Realistic Virtual Try-On for Diverse Fashion Items

Model ReleasesDGX agent

arXiv:2604.19748v1 Announce Type: new Abstract: Recent advances in image generation and editing have opened new opportunities for virtual try-on. However, existing methods still struggle to meet compl

Vibe coders are not going to like this. UC San Diego just published the first real field study of experienced developers using AI agents. Th…

AgentsDGX agent

Vibe coders are not going to like this. UC San Diego just published the first real field study of experienced developers using AI agents. They watched 13 of them code in the wild and surveyed 99 more.

Visual Adversarial Attack on Vision-Language Models for Autonomous Driving

SafetyDGX agent

arXiv:2411.18275v2 Announce Type: replace Abstract: Vision-language models (VLMs) have significantly advanced autonomous driving (AD) by enhancing reasoning capabilities. However, these models remain

Warmth and Competence in the Swarm: Designing Effective Human-Robot Teams

AgentsDGX agent

arXiv:2604.19270v1 Announce Type: new Abstract: As groups of robots increasingly collaborate with humans, understanding how humans perceive them is critical for designing effective human-robot teams.

We are excited to launch VideoGameBench on Antim Labs, created by @a1zhang, Thomas L. Griffiths (@cocosci_lab), @karthik_r_n, and @OfirPress…

AgentsDGX agent

VideoGameBench is a new benchmark launched on Antim Labs, created by a1zhang, Thomas L. Griffiths, Karthik R. N, and Ofir Press. The benchmark likely evaluates AI model performance on video game-relat

What’s new in the Agentic Data Cloud: Powering the System of Action

Model ReleasesDGX agent

Companies are shifting from gen AI that simply answers questions to autonomous agents that perceive, reason, and act on their behalf. Attempting to scale these agents on legacy stacks exposes structur

What’s next in Google AI infrastructure: Scaling for the agentic era

Model ReleasesDGX agent

AI is evolving from answering questions to reasoning and taking action. Companies who want to lead in today’s agentic era require computing infrastructure designed and optimized for these new requirem

21 Apr 2026

Advancing MAPF Toward the Real World: A Scalable Multi-Agent Realistic Testbed (SMART)

AgentsDGX agent

arXiv:2503.04798v3 Announce Type: replace Abstract: We present Scalable Multi-Agent Realistic Testbed (SMART), a realistic and efficient software tool for evaluating Multi-Agent Path Finding (MAPF) al

AI Insider: 'Adding a Human Makes Your Team Worse' Emad Mostaque | @EMostaque TIMESTAMPS : 00:00 The models too dangerous to release 10:06 W…

IndustryDGX agent

AI Insider: 'Adding a Human Makes Your Team Worse' Emad Mostaque | @EMostaque TIMESTAMPS : 00:00 The models too dangerous to release 10:06 Why physics needs axioms — and AI doesn't 11:43 The MIND fram

AIM 2025 Rip Current Segmentation (RipSeg) Challenge Report

Model ReleasesDGX agent

arXiv:2508.13401v3 Announce Type: replace Abstract: This report presents an overview of the AIM 2025 RipSeg Challenge, a competition designed to advance techniques for automatic rip current segmentati

Aligning Backchannel and Dialogue Context Representations via Contrastive LLM Fine-Tuning

SafetyDGX agent

arXiv:2604.16622v1 Announce Type: new Abstract: Backchannels (e.g., `yeah', `mhm', and `right') are short, non-interruptive feedback signals whose lexical form and prosody jointly convey pragmatic mea

AlphaContext: An Evolutionary Tree-based Psychometric Context Generator for Creativity Assessment

ApplicationsDGX agent

arXiv:2604.18398v1 Announce Type: new Abstract: Creativity has become a core competence in the era of LLMs and human-AI collaboration, underpinning innovation in real-world problem solving. Crucially,

Annotation-Assisted Learning of Treatment Policies From Multimodal Electronic Health Records

SafetyDGX agent

arXiv:2507.20993v3 Announce Type: replace Abstract: We study how to learn treatment policies from multimodal electronic health records (EHRs) that consist of tabular data and clinical text. These poli

Anthropic gets $5B investment from Amazon, will use it to buy Amazon chips

Model ReleasesDGX agent

Amazon announced a 5 billion investment in Anthropic, with up to 20 billion more tied to commercial milestones. Anthropic committed to spending over $100 billion on AWS technologies over the next deca

Automatic Dataset Construction (ADC): Sample Collection, Data Curation, and Beyond

Model ReleasesDGX agent

arXiv:2408.11338v2 Announce Type: replace-cross Abstract: Large-scale data collection is essential for developing personalized training data, mitigating the shortage of training data, and fine-tuning

Automatic Slide Updating with User-Defined Dynamic Templates and Natural Language Instructions

Model ReleasesDGX agent

arXiv:2604.17894v1 Announce Type: new Abstract: Presentation slides are a primary medium for data-driven reporting, yet keeping complex, analytics-style decks up to date remains labor-intensive. Exist

BenchMarker: An Education-Inspired Toolkit for Highlighting Flaws in Multiple-Choice Benchmarks

Model ReleasesDGX agent

arXiv:2602.06221v2 Announce Type: replace Abstract: Multiple-choice question answering (MCQA) is standard in NLP, but benchmarks lack rigorous quality control. We present BenchMarker, an education-ins

BIASEDTALES-ML: A Multilingual Dataset for Analyzing Narrative Attribute Distributions in LLM-Generated Stories

SafetyDGX agent

arXiv:2604.17008v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly used to generate narrative content, including children's stories, which play an important role in social a

CanonSLR: Canonical-View Guided Multi-View Continuous Sign Language Recognition

ApplicationsDGX agent

arXiv:2604.18184v1 Announce Type: new Abstract: Continuous Sign Language Recognition (CSLR) has achieved remarkable progress in recent years; however, most existing methods are developed under single-

CAPC-CG: A Large-Scale, Expert-Directed LLM-Annotated Corpus of Adaptive Policy Communication in China

SafetyDGX agent

arXiv:2510.08986v2 Announce Type: replace Abstract: We introduce CAPC-CG, the Chinese Adaptive Policy Communication (Central Government) Corpus, the first open dataset of Chinese policy directives ann

Capture Timing-Attention of Events in Clinical Time Series

Model ReleasesDGX agent

arXiv:2602.10385v2 Announce Type: replace Abstract: Automatically discovering personalized sequential events from large-scale time-series data is crucial for enabling precision medicine in clinical re

CaseFacts: A Benchmark for Legal Fact-Checking and Precedent Retrieval

Model ReleasesDGX agent

arXiv:2601.17230v2 Announce Type: replace Abstract: Automated Fact-Checking has largely focused on verifying general knowledge against static corpora, overlooking high-stakes domains like law where tr

CATP: Confidence-Aware Token Pruning for Camouflaged Object Detection

Model ReleasesDGX agent

arXiv:2604.16854v1 Announce Type: new Abstract: Camouflaged Object Detection (COD) aims to segment targets that share extreme textural and structural similarities with their complex environments. Leve

CBR-to-SQL: Rethinking Retrieval-based Text-to-SQL using Case-based Reasoning in the Healthcare Domain

ApplicationsDGX agent

arXiv:2603.05569v2 Announce Type: replace-cross Abstract: Extracting insights from Electronic Health Record (EHR) databases often requires SQL expertise, creating a barrier for clinical decision-makin

CFMS: Towards Explainable and Fine-Grained Chinese Multimodal Sarcasm Detection Benchmark

Model ReleasesDGX agent

arXiv:2604.16372v1 Announce Type: new Abstract: Multimodal sarcasm detection has recently garnered significant attention. However, existing benchmarks suffer from coarse-grained annotations and limite

CORP: A Multi-Modal Dataset for Campus-Oriented Roadside Perception Tasks

Model ReleasesDGX agent

arXiv:2404.03191v3 Announce Type: replace Abstract: Numerous roadside perception datasets have been introduced to propel advancements in autonomous driving and intelligent transportation systems resea

CRISP: Compressing Redundancy in Chain-of-Thought via Intrinsic Saliency Pruning

SafetyDGX agent

arXiv:2604.17297v1 Announce Type: new Abstract: Long Chain-of-Thought (CoT) reasoning is pivotal for the success of recent reasoning models but suffers from high computational overhead and latency. Wh

Decoding AI Tutor Effects for Educational Measurement: Temporal, Multi-Outcome, and Behavior-Cognitive Analysis

SafetyDGX agent

arXiv:2604.16366v1 Announce Type: cross Abstract: Artificial intelligence (AI) tutors have become increasingly popular in learning environments. In this study, we propose an AI agent prototype framewo

Differential Privacy in Two-Layer Networks: How DP-SGD Harms Fairness and Robustness

SafetyDGX agent

arXiv:2603.04881v2 Announce Type: replace Abstract: Differentially private learning is essential for training models on sensitive data, but empirical studies consistently show that it can degrade perf

DSH-Bench: A Difficulty- and Scenario-Aware Benchmark with Hierarchical Subject Taxonomy for Subject-Driven Text-to-Image Generation

Model ReleasesDGX agent

arXiv:2603.08090v2 Announce Type: replace Abstract: Significant progress has been achieved in subject-driven text-to-image (T2I) generation, which aims to synthesize new images depicting target subjec

Dual-Anchoring: Addressing State Drift in Vision-Language Navigation

AgentsDGX agent

arXiv:2604.17473v1 Announce Type: new Abstract: Vision-Language Navigation(VLN) requires an agent to navigate through 3D environments by following natural language instructions. While recent Video Lar

EgoWalk: A Multimodal Dataset for Robot Navigation in the Wild

ApplicationsDGX agent

arXiv:2505.21282v2 Announce Type: replace Abstract: Data-driven navigation algorithms are critically dependent on large-scale, high-quality real-world data collection for successful training and robus

EvoCoT: Overcoming the Exploration Bottleneck in Reinforcement Learning

Model ReleasesDGX agent

arXiv:2508.07809v5 Announce Type: replace Abstract: Reinforcement learning with verifiable reward (RLVR) has become a promising paradigm for post-training large language models (LLMs) to improve their

Expert-Annotated Embryo Image Dataset with Natural Language Descriptions for Evidence-Based Patient Communication in IVF

TutorialsDGX agent

arXiv:2604.16528v1 Announce Type: new Abstract: Embryo selection is one of multiple crucial steps in in-vitro fertilization, commonly based on morphological assessment by clinical embryologists. Altho

Fairness Constraints in High-Dimensional Generalized Linear Models

SafetyDGX agent

arXiv:2604.16610v1 Announce Type: cross Abstract: Machine learning models often inherit biases from historical data, raising critical concerns about fairness and accountability. Conventional fairness

← Previous
1…413414415416417…424
Next →