This is your work, valued
Senior Machine Learning Engineer @Roblox
BERT_for_ABSA. In this work (Targeted) Aspect-Based Sentiment Analysis task is converted to a sentence-pair classification task and a pre-trained BERT model is fine-tuned on it.
★ 32Keyframes-GAN. [IEEE TMM 2023] This is the official repo of the paper "Perceptual Quality Improvement in Videoconferencing using Keyframes-based GAN".
★ 17MeanShiftClustering. C++, OpenMP and CUDA implementation of Mean Shift clustering algorithm
★ 14Adversarial_attacks_defense. In this work the proposed defense strategy is evaluated against two black-box adversarial attacks, Hop Skip Jump and Square
★ 8Gutenberg. Gutenberg is a pipeline for training a neural network in segmenting and recognising frequent words in early printed books, in particular we focus on Gutenberg’s Bible.
★ 7TwitterRealTimeSentimentAnalysis. Lambda architecture implementation using Apache Storm, Hadoop and HBase to perform Twitter real-time sentiment analysis
★ 7Segway. We implement a control system that stabilize the robot in its vertical position, which corresponds to the unstable equilibrium state. The hardware board used to develop the robot is a LEGO MINDSTORM. We develop also an Android app for remote control.
★ 5IISA. [ICCV 2025] - Image Intrinsic Scale Assessment: Bridging the Gap Between Quality and Resolution
★ 3SocialDistancingDetector. Python
★ 3smartaccess. A web app for user access management
★ 2Depth-of-field. JavaScript
★ 2Depth-of-field-demo. HTML
★ 1Orbit. Svelte
★ 4IISA. [ICCV 2025] - Image Intrinsic Scale Assessment: Bridging the Gap Between Quality and Resolution
★ 17FRED. FRED: The Florence RGB-Event Drone Dataset
★ 82easymcp. TypeScript
★ 32torchmetrics. Machine learning metrics for distributed, scalable PyTorch applications.
★ 2.5kURAE. [ICML 2025] Official PyTorch implementation of paper "Ultra-Resolution Adaptation with Ease".
★ 118SAMWISE. [CVPR 2025 Highlight] "SAMWISE: Infusing Wisdom in SAM2 for Text-Driven Video Segmentation"
★ 386MDFS. [TMM-2024] Official Pytorch implementation of "Opinion-Unaware Blind Image Quality Assessment using Multi-Scale Deep Feature Statistics".
★ 5zero. [NeurIPS '24] Frustratingly easy Test-Time Adaptation of VLMs!!
★ 64Awesome-LLM-Anomaly-OOD-Detection. [NAACL 2025] Large Language Models for Anomaly and Out-of-Distribution Detection
★ 123Video-Quality-Assessment-A-Comprehensive-Survey. The Most Comprehensive Survey of Video Quality Assessment to Date.
★ 99Cross-the-Gap. [ICLR 2025] - Cross the Gap: Exposing the Intra-modal Misalignment in CLIP via Modality Inversion
★ 70EmoVOCA. [WACV 2025] - EmoVOCA: Speech-Driven Emotional 3D Talking Heads
★ 47OfflineMania. [COG24] - Official repository of "OfflineMania: A Benchmark Environment for Offline Reinforcement Learning in Racing Games"
★ 12awesome-machine-learning-startups. List of startups doing AI & ML
★ 438paperlists. Processed / Cleaned Data for Paper Copilot
★ 952DeepAtrialAnomalyDetection. [BSPC 2024] All-in-one electrical atrial substrate indicators with deep anomaly detection
★ 3IQA-Dataset. A unified interface for downloading and loading popular Image Quality Assessment (IQA) datasets.
★ 154WritingAIPaper. Writing AI Conference Papers: A Handbook for Beginners
★ 3.9kScanTalk. [ECCV 2024] - ScanTalk: 3D Talking Heads from Unregistered Scans
★ 55awesome-comics-understanding. The official repo of the Comics Survey: "A missing piece in Vision and Language: A Survey on Comics Understanding"
★ 139UIQA. Official Code for Assessing UHD Image Quality from Aesthetics, Distortions, and Saliency
★ 24LDRE. [SIGIR'2024 Best Paper Honorable Mention] Official repository for "LDRE: LLM-based Divergent Reasoning and Ensemble for Zero-Shot Composed Image Retrieval"
★ 105TOIS25-Awesome-Composed-Image-Retrieval. Collection of Composed Image Retrieval (CIR) papers.
★ 361AIMLInterviews. This repo is meant to serve as a guide for Machine Learning/AI technical interviews.
★ 8.7kbridge-score. [ECCV 2024] BRIDGE: Bridging Gaps in Image Captioning Evaluation with Stronger Visual Cues.
★ 13CoDE. [ECCV'24] Contrasting Deepfakes Diffusion via Contrastive Learning and Global-Local Similarities
★ 52safe-clip. Safe-CLIP: Removing NSFW Concepts from Vision-and-Language Models. ECCV 2024
★ 68ScanTalk. Python
★ 7SULAND-Dataset. Dataset for Surface Landmine detection. Videos are taken in Italy (Faculty of Engineering, Florence) and USA (Franklyn and Marshal college, Philadelphia).
★ 7KDPL. [ECCV 2024] - Improving Zero-shot Generalization of Learned Prompts via Unsupervised Knowledge Distillation
★ 62magiclens. [ICML'24 Oral] "MagicLens: Self-Supervised Image Retrieval with Open-Ended Instructions"
★ 211COVER. 🏆 [CVPRW 2024] COVER: A Comprehensive Video Quality Evaluator. 🥇 Winner solution for Video Quality Assessment Challenge at the 1st AIS 2024 workshop @ CVPR 2024
★ 100A-Bench. [ICLR 2025] What do we expect from LMMs as AIGI evaluators and how do they perform?
★ 123rscir. [IGARSS 2024 ORAL] Official implementation of "Composed Image Retrieval for Remote Sensing".
★ 86MambaOut. MambaOut: Do We Really Need Mamba for Vision? (CVPR 2025)
★ 2.7kcomposed-video-retrieval. Composed Video Retrieval
★ 62textual_inversion. Jupyter Notebook
★ 3.1kDassl.pytorch. A PyTorch toolbox for domain generalization, domain adaptation and semi-supervised learning.
★ 1.4kiamcl2r. [CVPR 2024 Highlight] - Stationary Representations: Optimally Approximating Compatibility and Implications for Improved Model Replacements (notable top 2.8%)
★ 13QualiCLIP. Quality-Aware Image-Text Alignment for Opinion-Unaware Image Quality Assessment
★ 132Awesome-Low-Level-Vision-Research-Groups. A Collection of Low Level Vision Research Groups
★ 246Vision_by_Language. [ICLR 2024] Official repository for "Vision-by-Language for Training-Free Compositional Image Retrieval"
★ 89Co-Instruct. ④[ECCV 2024 Oral, Comparison among Multiple Images!] A study on open-ended multi-image quality comparison: a dataset, a model and a benchmark.
★ 87TReS. No-Reference Image Quality Assessment via Transformers, Relative Ranking, and Self-Consistency
★ 161iqa-ContentSep. Official pytorch implementation of the WACV'23 paper "No Reference Opinion Unaware Quality Assessment of Authentically Distorted Images".
★ 3InstructIR. [ECCV 2024] InstructIR: High-Quality Image Restoration Following Human Instructions https://huggingface.co/spaces/marcosv/InstructIR
★ 742simpool. [ICCV 2023] Official implementation of "Keep It SimPool: Who Said Supervised Transformers Suffer from Attention Deficit?".
★ 102SPRC. 【ICLR 2024, Spotlight】Sentence-level Prompts Benefit Composed Image Retrieval
★ 94GRepQ. Official repository for our paper titled "Learning Generalizable Perceptual Representations for Data-Efficient No-Reference Image Quality Assessment".
★ 16WACV-2024-Papers. WACV 2024 Papers: Discover cutting-edge research from WACV 2024, the leading computer vision conference. Stay updated on the latest in computer vision and deep learning, with code included. ⭐ support visual intelligence development!
★ 97Land-Diffuser. The Land-Diffuser is a novel application of the Denoising Diffusion Probabilistic Model (DDPM) in the realm of 3D Talking Head generation from raw audio inputs.
★ 13SPAQ. [CVPR'20] Official SPAQ & Implementation
★ 198dotmap. Dot access dictionary with dynamic hierarchy creation and ordered iteration
★ 481lincir. Official Pytorch implementation of LinCIR: Language-only Training of Zero-shot Composed Image Retrieval (CVPR 2024)
★ 148Swin-Transformer. This is an official implementation for "Swin Transformer: Hierarchical Vision Transformer using Shifted Windows".
★ 16kFriendsDontLetFriends. Friends don't let friends make certain types of data visualization - What are they and why are they bad.
★ 7.1kfashion-iq. Python
★ 171genecis. Code and Models for "GeneCIS A Benchmark for General Conditional Image Similarity"
★ 61HWD. Python
★ 27ARNIQA. [WACV 2024 Oral] - ARNIQA: Learning Distortion Manifold for Image Quality Assessment
★ 156TAPE. [WACV 2024] - Reference-based Restoration of Digitized Analog Videotapes
★ 60CLIP-LIT. [ICCV 2023, Oral] Iterative Prompt Learning for Unsupervised Backlit Image Enhancement
★ 351vic. Code implementation of our NeurIPS 2023 paper: Vocabulary-free Image Classification
★ 107daclip-uir. [ICLR 2024] Controlling Vision-Language Models for Universal Image Restoration. 5th place in the NTIRE 2024 Restore Any Image Model in the Wild Challenge.
★ 816PIQ2023. Jupyter Notebook
★ 108Awesome-Aesthetic-Evaluation-and-Cropping.
★ 409Q-Bench. ①[ICLR2024 Spotlight] (GPT-4V/Gemini-Pro/Qwen-VL-Plus+16 OS MLLMs) A benchmark for multi-modality LLMs (MLLMs) on low-level vision and visual quality assessment.
★ 287relvit. Official code of "Where are my Neighbors? Exploiting Patches Relations in Self-Supervised Vision Transformer", Guglielmo Camporese, Elena Izzo, Lamberto Ballan. BMVC, 2022.
★ 21doc2graph. Doc2Graph transforms documents into graphs and exploit a GNN to solve several tasks.
★ 139RankIQA. The rep for the RankIQA paper in ICCV 2017
★ 451unifree. Python
★ 1.4kAwesome-Image-Quality-Assessment. A comprehensive collection of IQA papers
★ 1.5kISSUES. [ICCVW 2023] - Mapping Memes to Words for Multimodal Hateful Meme Classification
★ 27CLIP-IQA. [AAAI 2023] Exploring CLIP for Assessing the Look and Feel of Images
★ 492chatgpt-prompts-for-academic-writing. This list of writing prompts covers a range of topics and tasks, including brainstorming research ideas, improving language and style, conducting literature reviews, and developing research plans.
★ 4.9kdysgraphia-detection. Help building intelligent systems in support of experts to recognise dysgraphia in young subjects, using Machine Learning and Smart Pen
★ 5Diff-JPEG. Official and maintained implementation of the paper "Differentiable JPEG: The Devil is in the Details" (WACV 2024)
★ 114ICCV-2023-25-Papers. ICCV 2023-2025 Papers: Discover cutting-edge research from ICCV 2023-25, the leading computer vision conference. Stay updated on the latest in computer vision and deep learning, with code included. ⭐ support visual intelligence development!
★ 968al-folio. A beautiful, simple, clean, and responsive Jekyll theme for academics
★ 16kCoVR. Official PyTorch implementation of the paper "CoVR: Learning Composed Video Retrieval from Web Video Captions".
★ 119Real-ESRGAN. Real-ESRGAN aims at developing Practical Algorithms for General Image/Video Restoration.
★ 36kStableVideo. [ICCV 2023] StableVideo: Text-driven Consistency-aware Diffusion Video Editing
★ 1.4kKAIR. Image Restoration Toolbox (PyTorch). Training and testing codes for DPIR, USRNet, DnCNN, FFDNet, SRMD, DPSR, BSRGAN, SwinIR
★ 3.5kIQA-PyTorch. 🔎 🖼️ 🔥PyTorch Toolbox for Image Quality Assessment, including PSNR, SSIM, LPIPS, FID, NIQE, NRQM(Ma), MUSIQ, TOPIQ, NIMA, DBCNN, BRISQUE, PI and more...
★ 3.3kImage-Color-Aesthetics-and-Quality-Assessment. 🔥[ICCV 2023, Official Code] for paper "Thinking Image Color Aesthetics Assessment: Models, Datasets and Benchmarks". Official Weights and Demos provided. 首个面向图像色彩主观美学评估的数据集、算法和benchmark.
★ 214ICCV2025-Papers-with-Code. ICCV 2025 论文和开源项目合集
★ 2.9kCoOp. Prompt Learning for Vision-Language Models (IJCV'22, CVPR'22)
★ 2DragGAN. Official Code for DragGAN (SIGGRAPH 2023)
★ 36kijepa. Official codebase for I-JEPA, the Image-based Joint-Embedding Predictive Architecture. First outlined in the CVPR paper, "Self-supervised learning from images with a joint-embedding predictive architecture."
★ 3.5kbig_vision. Official codebase used to develop Vision Transformer, SigLIP, MLP-Mixer, LiT and more.
★ 3.5ks2l-s2d. [ICIAP 2023] Learning Landmarks Motion from Speech for Speaker-Agnostic 3D Talking Heads Generation
★ 59CompoDiff. Official Pytorch implementation of "CompoDiff: Versatile Composed Image Retrieval With Latent Diffusion" (TMLR 2024)
★ 88cores-compatibility. This repo is the official implementation of "CoReS: Compatible Representations via Stationarity" Niccolò Biondi, Federico Pernici, Matteo Bruni, and Alberto Del Bimbo, IEEE TPAMI
★ 4ladi-vton. [ACM MM 2023] - LaDI-VTON: Latent Diffusion Textual-Inversion Enhanced Virtual Try-On
★ 466LSDIR.
★ 92NEFER. [CVPR 2023] NEFER a Dataset for Neuromorphic Event-based Facial Expression Recognition
★ 30dress-code. Dress Code: High-Resolution Multi-Category Virtual Try-On. ECCV 2022
★ 658multimodal-garment-designer. This is the official repository for the paper "Multimodal Garment Designer: Human-Centric Latent Diffusion Models for Fashion Image Editing". ICCV 2023
★ 445CIRR. Official repository of ICCV 2021 - Image Retrieval on Real-life Images with Pre-trained Vision-and-Language Models
★ 135analog-video-restoration. [ACM MM 2022 - Demo] Restoration of Analog Videos Using Swin-UNet
★ 38cores-compatibility. This repo is the official implementation of "CoReS: Compatible Representations via Stationarity" Niccolò Biondi, Federico Pernici, Matteo Bruni, and Alberto Del Bimbo, IEEE TPAMI
★ 4SEARLE. [ICCV 2023] - Zero-shot Composed Image Retrieval with Textual Inversion
★ 198CIRCO. [ICCV 2023] - Composed Image Retrieval on Common Objects in context (CIRCO) dataset
★ 87General-GPT. Jupyter Notebook
★ 65composed_image_retrieval. Shell
★ 197Synthetic-Social-Agents-Dataset. Python
★ 3SwinIR. SwinIR: Image Restoration Using Swin Transformer (official repository)
★ 5.6kpytorch-image-models. The largest collection of PyTorch image encoders / backbones. Including train, eval, inference, export scripts, and pretrained weights -- ResNet, ResNeXT, EfficientNet, NFNet, Vision Transformer (ViT), MobileNetV4, MobileNet-V3 & V2, RegNet, DPN, CSPNet, Swin Transformer, MaxViT, CoAtNet, ConvNeXt, and more
★ 37kESA. [ECCV2022] Code for Explainable Sparse Attention for Memory-based Trajectory Predictors
★ 7Modality-Gap. Mind the Gap: Understanding the Modality Gap in Multi-modal Contrastive Representation Learning
★ 178CompatibleLifelongRepresentation. This repo is the official implementation of "CL^2R: Compatible Lifelong Learning Representations" Biondi et al. 2022 ACM TOMM
★ 9CLIP4Cir. [ACM TOMM 2023] - Composed Image Retrieval using Contrastive Learning and Task-oriented CLIP-based Features
★ 195stable-diffusion. A latent text-to-image diffusion model
★ 73kdropzone. Dropzone is an easy to use drag'n'drop library. It supports image previews and shows nice progress bars.
★ 18kCCPL. [ECCV 2022 Oral] Official Pytorch implementation of CCPL and SCTNet
★ 206MANTRA-CVPR20. Official Pytorch code for MANTRA - Memory Augmented Neural Trajectory Predictor (CVPR2020)
★ 78ContrastiveSupervisedDistillation. This repo contains the code of "Contrastive Supervised Distillation for Continual Representation Learning", Tommaso Barletti, Niccolò Biondi, Federico Pernici, Matteo Bruni, and Alberto Del Bimbo, ICIAP2021
★ 20Mean_Shift_clustering. C++
★ 2Kmeans. Cuda
★ 3CLIP4CirDemo. [CVPR 2022 - Demo Track] - Effective conditioned and composed image retrieval combining CLIP-based features
★ 85awesome-italia-remote. A list of remote-friendly or full-remote companies that targets Italian talents.
★ 2.6kwordle-it. Italian version of Wordle
★ 141Citation_Intent_Classification. Citation Intent Classification in scientific papers using the Scicite dataset an Pytorch
★ 7Moving-Least-Squares. Numpy & PyTorch implementation of three algorithms of image deformation using moving least squares. http://dl.acm.org/citation.cfm?doid=1179352.1141920
★ 364MLQuestions. Machine Learning and Computer Vision Engineer - Technical Interview Questions
★ 4.8kaioway. AI on the way. An auto deep learning pipe dream. An RDBMS approach to deep learning. Declarative, explainable, scalable, optimizable, easy to deploy, all that good stuff.
★ 1.8kmachine-learning-interview. Machine Learning Interviews from FAANG, Snapchat, LinkedIn. I have offers from Snapchat, Coupang, Stitchfix etc. Blog: mlengineer.io.
★ 13kKobayashi-Colors. CNN for outfits classification, based on Kobayashi classes
★ 2CLIP. CLIP (Contrastive Language-Image Pretraining), Predict the most relevant text snippet given an image
★ 34kunifi-webex-dl. Download recorded lessons from unifi webex platform passing by the Moodle platform.
★ 16in-exhibition-app. Web-app with Vue.js for the IN Exhibition experience
★ 2AugmentBrain. In AugmentBrain we investigate the performance of different data augmentation methods for the classification of Motor Imagery (MI) data using a Convolutional Neural Network tailored for EEG named EEGNet.
★ 22BasketTracking. Basketball 🏀 action tracking and understanding using classical computer vision approaches and deep learning.
★ 75RoboLeague. A car soccer environment inspired by Rocket League for deep reinforcement learning experiments in an adversarial self-play setting.
★ 309robusta. Easy interop between Rust and Java
★ 371SupeRAuGAN. In SupeRAuGAN we implement a novel data augmentation technique tailored to Generative Adversarial Networks in order to reduce discriminator overfitting and stabilize training
★ 4fires-regression-analysis. Analysis of forest fires predictors in R
★ 2MT-BERT. In MT-BERT we reproduce a neural language understanding model which implements a Multi-Task Deep Neural Network (MT-DNN) for learning representations across multiple NLU tasks.
★ 22ABSA-BERT-pair. Utilizing BERT for Aspect-Based Sentiment Analysis via Constructing Auxiliary Sentence (NAACL 2019)
★ 519BrainPad. Classification of EEG signals from the brain 🧠 through OpenBCI hardware and Tensorflow-Keras API
★ 35facestretch. In facestretch we describe how we exploited the dlib’s facial landmarks in order to measure face deformation and perform an expression recognition task. We implemented multiple approaches, based metric learning, neural networks and geodesic distances
★ 5phone_finder. A simple program to look through Android phones' specs and generate an HTML report
★ 6dawntime. https://reddeadrecovery.gitlab.io/dawntime/ In dawntime we implement a volumetric light scattering effect based on the postprocessing technique described by Kenny Mitchell.
★ 8Segway. We implement a control system that stabilize the robot in its vertical position, which corresponds to the unstable equilibrium state. The hardware board used to develop the robot is a LEGO MINDSTORM. We develop also an Android app for remote control.
★ 5Appunti. Appunti per alcuni corsi
★ 8COVID-19. Project for COVID-19 pandemy
★ 8COVID-19. COVID-19 Italia - Monitoraggio situazione
★ 3.8ksmartaccess. A web app for user access management
★ 2