Skip to content
View xujustinj's full-sized avatar

Highlights

  • Pro

Organizations

@VectorInstitute @loolabs

Block or report xujustinj

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Make Zotero effective for us LaTeX holdouts

TypeScript 7,016 386 Updated Aug 12, 2026

Awesome papers about generative Information Extraction (IE) using Large Language Models (LLMs)

1,061 61 Updated Nov 18, 2024

CrossRE: A Cross-Domain Dataset for Relation Extraction (Findings of EMNLP 2022)

Python 49 Updated Aug 20, 2024

It was a simple bash script for installing and configuring Cursor on Debian/Ubuntu-based Linux distributions at a time when the DEB package was not released.

Shell 54 5 Updated Aug 27, 2025

BioCreative VI — Track 5: text mining chemical–protein interactions

Python 10 5 Updated Aug 21, 2021

🧙 Automates the installation and updating of the Cursor .AppImage for Linux users, resolving common issues during setup and effortlessly handling configurations, updates, and related tasks.

Shell 422 66 Updated Jan 21, 2025

Efficient LLM inference on Slurm clusters.

Python 106 14 Updated Aug 10, 2026

A playbook for systematically maximizing the performance of deep learning models.

30,280 2,420 Updated Jun 18, 2024

🍽️ Annotations for the public release of the EPIC-KITCHENS-100 dataset

Python 173 34 Updated Aug 1, 2022

[NeurIPS 2022] Egocentric Video-Language Pretraining

Python 261 24 Updated May 9, 2024
TypeScript 1 Updated May 29, 2025

One million English sentences, each split into two sentences that together preserve the original meaning, extracted from Wikipedia edits.

125 5 Updated Jun 3, 2019

This is a preprocessor/data-cleaner for the WebNLG dataset.

Python 10 3 Updated Oct 17, 2023

A Python library for processing and filtering TabLib

Python 14 4 Updated Aug 24, 2024

Type-safe FFmpeg bindings for Python & TypeScript — filters, typing, and docs

Python 1,160 22 Updated Aug 13, 2026

Code for benchmarking VLMs on zero and few-shot activity recognition

HTML 8 3 Updated Nov 27, 2024

An Easy-to-use, Scalable and High-performance Agentic RL Framework based on Ray (PPO & DAPO & REINFORCE++ & VLM & TIS & vLLM & Ray & Async RL)

Python 9,907 998 Updated Aug 13, 2026

An open source implementation of CLIP.

Python 14,060 1,301 Updated Aug 10, 2026

EMNLP2020 findings paper: Minimize Exposure Bias of Seq2Seq Models in Joint Entity and Relation Extraction

Python 51 7 Updated Dec 8, 2022

A dataset for multi-object multi-actor activity parsing

Jupyter Notebook 45 6 Updated Sep 29, 2023

Code for paper "Document-Level Argument Extraction by Conditional Generation". NAACL 21'

HTML 121 30 Updated Mar 28, 2023

Completing the Puzzle of All-in-One Event Understanding Benchmark with Event Arguments

Python 14 Updated Mar 12, 2024

Source code and dataset for EMNLP 2022 paper "MAVEN-ERE: A Unified Large-scale Dataset for Event Coreference, Temporal, Causal, and Subevent Relation Extraction".

Python 92 11 Updated Aug 26, 2023

Source code and dataset for EMNLP 2020 paper "MAVEN: A Massive General Domain Event Detection Dataset".

Python 168 38 Updated Jan 5, 2022

Universal memory layer for AI Agents

Python 63,168 7,369 Updated Aug 13, 2026

A collection of large question answering datasets

439 45 Updated Jul 1, 2024
Python 598 86 Updated Apr 26, 2021

Knowledge Graph for Legal Documents using Litigation Releases from the SEC website. Classifies into different crimes, extracts relevant information (violator, violation, action taken by authorities…

Jupyter Notebook 84 18 Updated Feb 15, 2022

SOIRP

Python 2 Updated Dec 13, 2023
Next