-
Luleå University of Technology
- Sweden
- https://www.linkedin.com/in/konstantina-nikolaidou-18009616b/
- @koni_nik
Stars
This repo contains a curated list of research papers and resources focusing on Handwritten Text Generation (HTG)
[NeurIPS 2025] Official implementation of "Instance-Level Composed Image Retrieval".
[ICCV 2025] Official repository of the paper "Talking to DINO: Bridging Self-Supervised Vision Backbones with Language for Open-Vocabulary Segmentation"
[ICCV 2025] What Changed? Detecting and Evaluating Instruction-Guided Image Edits with Multimodal Large Language Models
AskAnything in Charts - Powered by Qwen 2.5 7B model with LORA fine-tuned version for chart understanding!
Möbius Perspective Distortion (MPD) data-augmentation
ASTrA - Adversarial-Self-Supervised-Training-with-Adaptive-Attacks, ICLR 2025
Efficient On-device Budgeting for Differentially-Private Ad-Measurement Systems (SOSP '24)
Official PyTorch implementation of the WACV 2025 Oral paper "Composed Image Retrieval for Training-FREE DOMain Conversion".
[ECCV 2024] Official repo for UDiffText: A Unified Framework for High-quality Text Synthesis in Arbitrary Images via Character-aware Diffusion Models
Comics Dataset Framework for Comics Understanding
Perceptual Grouping in Contrastive Vision-Language Models (ICCV'23)
[2024-NeurIPS] TextCtrl: Diffusion-based Scene Text Editing with Prior Guidance Control
drogozhang / magiclens
Forked from google-deepmind/magiclens[ICML'24 Oral] "MagicLens: Self-Supervised Image Retrieval with Open-Ended Instructions"
Implementation of Autoregressive Diffusion in Pytorch
[ICLR2025] Halton Scheduler for Masked Generative Image Transformer
Official Implementation of the CrossMAE paper: Rethinking Patch Dependence for Masked Autoencoders
This repo hosts the code and models of "Masked Autoencoders that Listen".
Implementation of Vision Transformer, a simple way to achieve SOTA in vision classification with only a single transformer encoder, in Pytorch
LCM: Log Conformal Maps for Robust Representation Learning to Mitigate Perspective Distortion, ACCV 2024
A collection of literature after or concurrent with Masked Autoencoder (MAE) (Kaiming He el al.).
Generative Models by Stability AI
Implementation of the paper: "BRAVE : Broadening the visual encoding of vision-language models"
Handwritten Text Recognition and Character Detection
Prompt Learning for Vision-Language Models (IJCV'22, CVPR'22)
Official PyTorch implementation for "Merging and Splitting Diffusion Paths for Semantically Coherent Panoramas", presenting the Merge-Attend-Diffuse operator (ECCV24)
Official PyTorch Implementation of "Rethinking HTG Evaluation: Bridging Generation and Recognition" (Oral) - 1st Workshop on Critical Evaluation of Generative Models and their Impact on Society - E…