Highlights
- Pro
Stars
Official repo for the paper "Scaling Synthetic Data Creation with 1,000,000,000 Personas"
主要记录大语言大模型(LLMs) 算法(应用)工程师相关的知识及面试题
The agent harness performance optimization system. Skills, instincts, memory, security, and research-first development for Claude Code, Codex, Opencode, Cursor and beyond.
Curated visual catalog of 155+ vision-language model (VLM/MLLM) architectures: papers, diagrams, training recipes, datasets, and a release timeline for multimodal AI agents.
A python experiment in heuristic search algorithms – finding the shortest path from one wikipedia page to the other.
Guide to interviewing for industry machine learning roles (data/applied/research scientist, ML engineer, etc).
Machine Theory of Mind Reading List. Built upon EMNLP Findings 2023 Paper: Towards A Holistic Landscape of Situated Theory of Mind in Large Language Models
This repository contains code for the paper Gendered Mental Health Stigma in Masked Language Models.
A list of ethics related resources for researchers and practitioners of Natural Language Processing and Computational Linguistics
Tensors and Dynamic neural networks in Python with strong GPU acceleration
convolutional lstm implementation in pytorch
A logical, reasonably standardized, but flexible project structure for doing and sharing data science work.
An evolving list of electronic media data sets used to model mental-health status.
Transformers for Information Retrieval, Text Classification, NER, QA, Language Modelling, Language Generation, T5, Multi-Modal, and Conversational AI
Code and datasets for the paper "Humor Detection: A Transformer Gets the Last Laugh"