Federated Learning under Distributed Concept Drift (FedDrift)
This repository is the source code for our paper: Federated Learning under Distributed Concept Drift (AISTATS’23) (opens in new tab).
Discover an index of datasets, SDKs, APIs and open-source tools developed by Microsoft researchers and shared with the global academic community below. These experimental technologies—available through Azure AI Foundry Labs (opens in new tab)—offer a glimpse into the future of AI innovation.
This repository is the source code for our paper: Federated Learning under Distributed Concept Drift (AISTATS’23) (opens in new tab).
This is a benchmark tool to evaluate natural language-based network management using LLM-generated code.
This dataset serves as a benchmark for evaluting the performance and efficiency of anomaly detectors in east-west data center network traffic.
AImagery is an AI-powered multisensory relaxation system designed to reduce anxiety by providing personalized immersive experiences. The system uses AI to create guided imagery based on individual preferences and physiological feedback, incorporating elements like auditory…
VeriSMo: A formally verified security module for AMD confidential VMs.
ORCAS is a click-based dataset associated with the TREC Deep Learning Track. It covers 1.4 million of the TREC DL documents, providing 18 million connections to 10 million distinct queries.
Tip-of-the-tongue (ToT) known-item retrieval is defined as “an item identification task in which the searcher has previously experienced an item but cannot recall a reliable identifier” (i.e., “It’s on the tip of my tongue…”). The…
The TREC Deep Learning Track studies information retrieval in a large training data regime. This is the case where the number of training queries with at least one positive label is at least in the…
GitHub Publication Publication Publication Publication Publication
The Phi-3-Mini-128K-Instruct is a 3.8B parameters, lightweight, state-of-the-art open model trained with the Phi-3 datasets that includes both synthetic data and the filtered publicly available websites data with a focus on high-quality and reasoning dense…