A Sanity Check on Composed Image Retrieval
arXiv:2604.12904v1 Announce Type: new Abstract: Composed Image Retrieval (CIR) aims to retrieve a target image based on a query composed of a reference image, and a relative caption that specifies the
Knowledge catalogue
arXiv:2604.12904v1 Announce Type: new Abstract: Composed Image Retrieval (CIR) aims to retrieve a target image based on a query composed of a reference image, and a relative caption that specifies the
arXiv:2604.11944v1 Announce Type: new Abstract: Diabetes devices, including Continuous Glucose Monitoring (CGM), Smart Insulin Pens, and Automated Insulin Delivery systems, generate rich time-series d
arXiv:2603.18104v3 Announce Type: replace Abstract: Prevailing AI training infrastructure assumes reverse-mode automatic differentiation over IEEE-754 arithmetic. The memory overhead of training relat
arXiv:2604.12999v1 Announce Type: new Abstract: We introduce HypoExplore, an agentic framework that formulates neural architecture discovery for visual recognition as a hypothesis-driven scientific in
arXiv:2603.21011v2 Announce Type: replace-cross Abstract: Finite element (FE) analysis guides the design and verification of nearly all manufactured objects. It is at the core of computational enginee
This r/ChatGPT thread explores the ongoing debate over whether AI hallucinations — where large language models confidently generate false or fabricated information — represent an inherent flaw or an u
Six-month-old cybersecurity startup Artemis Global Technologies Inc. today disclosed that it has raised 70 million in funding. The capital arrived in two tranches. Felicis led the largest of the two t
arXiv:2604.12857v1 Announce Type: new Abstract: Autonomous vehicles (AVs) are now operating on public roads, which makes their testing and validation more critical than ever. Simulation offers a safe
Cybersecurity asset management startup Axonius Inc. today unveiled three expansions to its Asset Cloud platform that see the addition of artificial intelligence-driven remediation, extended coverage t
arXiv:2604.12191v1 Announce Type: new Abstract: Current evaluations of large language models aggregate performance across diverse tasks into single scores. This obscures fine-grained ability variation
arXiv:2604.12686v1 Announce Type: cross Abstract: Recent advances in deep learning underscore the need for systems that can not only acquire new knowledge through Continual Learning (CL) but also remo
arXiv:2604.12325v1 Announce Type: cross Abstract: We consider the problem of offline black-box optimization, where the goal is to discover optimal designs (e.g., molecules or materials) from past expe
This r/ChatGPT thread discusses how job seekers are leveraging ChatGPT throughout the job search process, including tailoring resumes and cover letters to specific job descriptions, preparing for inte
arXiv:2604.13022v1 Announce Type: cross Abstract: The Energy Conserving Descent (ECD) algorithm was recently proposed (De Luca & Silverstein, 2022) as a global non-convex optimization method. Unlike g
arXiv:2604.05821v2 Announce Type: replace Abstract: Existing multilingual embedding models often encounter challenges in cross-lingual scenarios due to imbalanced linguistic resources and less conside
arXiv:2604.12308v1 Announce Type: new Abstract: Individuals' concerns about data privacy and AI safety are highly contextualized and extend beyond sensitive patterns. Addressing these issues requires
How to join LLM traces with billing, infrastructure, and customer data using Iceberg and BigQuery If you run AI agents in production, you’ve probably run into a simple problem: you... The post Data Fa
arXiv:2511.08439v2 Announce Type: replace Abstract: Dataset integrity is fundamental to the safety and reliability of AI systems, especially in autonomous driving. This paper presents a structured fra
arXiv:2604.12260v1 Announce Type: new Abstract: We study decentralized learning over networks where data are distributed across nodes without a central coordinator. Random walk learning is a token-bas
arXiv:2604.12615v1 Announce Type: new Abstract: This report summarizes the results of the first edition of the Large Language Model (LLM) Testing competition, held as part of the DeepTest workshop at
arXiv:2604.12293v1 Announce Type: new Abstract: As the number of fatalities involving Autonomous Vehicles increase, the need for a universal method of communicating between vehicles and other agents o
arXiv:2604.01315v2 Announce Type: replace Abstract: Money launderers take advantage of limitations in existing detection approaches by hiding their financial footprints in a deceitful manner. They man
arXiv:2604.12343v1 Announce Type: new Abstract: We address the challenging task of detecting the precise moment when hands make contact with objects in egocentric videos. This frame-level detection is
arXiv:2604.12778v1 Announce Type: cross Abstract: Purpose: Accurate dose calculation is essential in radiotherapy for precise tumor irradiation while sparing healthy tissue. With the growing adoption
arXiv:2604.12320v1 Announce Type: cross Abstract: While video large language models (Video-LLMs) excel in understanding slow-paced, real-world egocentric videos, their capabilities in high-velocity, i
arXiv:2604.11932v1 Announce Type: new Abstract: Solving pattern recognition problems using imbalanced databases is a hot topic, which entices researchers to bring it into focus. Therefore, we consider
arXiv:2604.12968v1 Announce Type: cross Abstract: Balancing convergence speed, generalization capability, and computational efficiency remains a core challenge in deep learning optimization. First-ord
Netgear became the first retail consumer router company to receive a conditional FCC exemption from the agency's ban on foreign-made routers, allowing it to introduce new models and push software upda
arXiv:2503.05167v3 Announce Type: replace Abstract: Traditional Chinese medicine (TCM) exhibits remarkable therapeutic efficacy in healthcare through patient-specific formulas. However, current AI-bas
Google DeepMind's Gemini 3.1 Flash TTS is a text-to-speech model delivering improved controllability, expressivity, and quality for developers, enterprises, and everyday users building AI-speech appli
A community thread on r/MachineLearning where users are invited to share research ideas, project concepts, or suggestions related to machine learning. The '[N]' tag indicates it is a discussion post r
arXiv:2604.12442v1 Announce Type: new Abstract: In derivational morphology, what mechanisms govern the variation in form-meaning relations between words? The answers to this type of questions are typi
Google launched its 'Google app for desktop' for Windows 10 and 11 users globally on April 14, 2026, bringing Google Search, Drive, files, and apps directly to the desktop via a keyboard shortcut (Alt
Moss is a YC-backed high-performance runtime for real-time semantic search that delivers sub-10ms lookups, instant index updates, and zero infrastructure overhead, running where the agent lives — clou
I'm going all in on Hermes (@NousResearch, @Teknium1) as my entire agent and coding stack. Six profiles. One shared self-hosted memory store. Zero hosted-coder dependencies. The fleet: - pmax-mousa —
arXiv:2604.12805v1 Announce Type: new Abstract: Image-to-image translation (I2I) is a fundamental task in computer vision, focused on mapping an input image from a source domain to a corresponding ima
arXiv:2507.13647v2 Announce Type: replace-cross Abstract: Real-time trajectory planning for unmanned aerial vehicles (UAVs) in dynamic environments remains a key challenge due to high computational de
arXiv:2604.11827v1 Announce Type: cross Abstract: Machine learning is revolutionizing chemistry. Beyond the value of predictive models accelerating virtual screening, generative AI aims at enabling in
arXiv:2510.23538v2 Announce Type: replace Abstract: The scope of neural code intelligence is rapidly expanding beyond text-based source code to encompass the rich visual outputs that programs generate
arXiv:2604.12596v1 Announce Type: cross Abstract: We introduce KumoRFM-2, the next iteration of a pre-trained foundation model for relational data. KumoRFM-2 supports in-context learning as well as fi
arXiv:2509.10026v4 Announce Type: replace Abstract: As large vision language models (VLMs) advance, their capabilities in multilingual visual question answering (mVQA) have significantly improved. Cha
arXiv:2511.17714v5 Announce Type: replace Abstract: Standard decision frameworks address uncertainty about facts but assume fixed options and values. We extend the Jeffrey-Bolker framework to model re
arXiv:2604.11995v1 Announce Type: new Abstract: The central goal of active learning is to gather data that maximises downstream predictive performance, but popular approaches have limited flexibility
Lyra 2.0 is an NVIDIA research project that extends the original Lyra framework for generative 3D world creation, using a self-distillation approach to distill the implicit 3D knowledge in video diffu
arXiv:2604.12917v1 Announce Type: new Abstract: Image restoration under adverse conditions, such as underwater, haze or fog, and low-light environments, remains a highly challenging problem due to com
arXiv:2604.12416v1 Announce Type: cross Abstract: In this review I summarize how machine learning can be used in lattice gauge theory simulations and what ap-proaches are currently available to improv
arXiv:2603.19796v3 Announce Type: replace-cross Abstract: Binary on/off thrusters are commonly used for spacecraft attitude and position control during proximity operations. However, their discrete na
arXiv:2507.04227v2 Announce Type: replace-cross Abstract: Recent years have witnessed a rapid development of mobile GUI agents powered by large language models (LLMs), which can autonomously execute d
arXiv:2604.12766v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) typically relies on a flat retrieval paradigm that maps queries directly to static, isolated text segments. This ap
arXiv:2604.12668v1 Announce Type: new Abstract: The Diffusion Probabilistic Model (DPM) achieves remarkable performance in image generation, while its increasing parameter size and computational overh
arXiv:2604.12356v1 Announce Type: new Abstract: Accurate estimation of food nutrition plays a vital role in promoting healthy dietary habits and personalized diet management. Most existing food datase
arXiv:2604.12075v1 Announce Type: cross Abstract: The tumor microenvironment (TME) plays a central role in cancer progression, treatment response, and patient outcomes, yet large-scale, consistent, an
arXiv:2604.12872v1 Announce Type: new Abstract: Object Goal Navigation (ObjectNav) refers to an agent navigating to an object in an unseen environment, which is an ability often required in the accomp
arXiv:2604.12986v1 Announce Type: cross Abstract: Autonomous AI agents are rapidly transitioning from experimental tools to operational infrastructure, with projections that 80% of enterprise applicat
arXiv:2008.07644v3 Announce Type: replace-cross Abstract: Jigsaw puzzle solving, the problem of constructing a coherent whole from a set of non-overlapping unordered visual fragments, is fundamental t
arXiv:2604.12113v1 Announce Type: cross Abstract: Visual Foundation Models (VFMs) such as the Segment Anything Model (SAM) have significantly advanced broad use of image segmentation. However, SAM and
arXiv:2604.12970v1 Announce Type: cross Abstract: Multimodal federated learning enables privacy-preserving collaborative model training across healthcare institutions. However, a fundamental challenge
arXiv:2604.11943v1 Announce Type: cross Abstract: An OS kernel that runs LLM inference internally can read logit distributions before any text is generated -- and act on them as a governance primitive
arXiv:2604.12160v1 Announce Type: new Abstract: Reasoning post-training with reinforcement learning from verifiable rewards (RLVR) is typically studied in centralized settings, yet many realistic appl
arXiv:2511.10453v3 Announce Type: replace-cross Abstract: Large language models often respond to ambiguous requests by implicitly committing to one interpretation, frustrating users and creating safet