-
NTT, Inc.
- Tokyo
-
12:33
(UTC +09:00) - in/daisuke-kikuta-b454661a2
- https://scholar.google.com/citations?user=0jYdmCMAAAAJ
Stars
The fastest, litest AI Gateway. Rust core with Python SDK. Call 100+ LLM APIs in OpenAI (or native) format with cost tracking, guardrails, load balancing, and logging [Bedrock, Azure, OpenAI, Anthr…
Playwright is a framework for Web Testing and Automation. It allows testing Chromium, Firefox and WebKit with a single API.
Multi-SWE-bench: A Multilingual Benchmark for Issue Resolving
Claude Code is an agentic coding tool that lives in your terminal, understands your codebase, and helps you code faster by executing routine tasks, explaining complex code, and handling git workflo…
Python tool for converting files and office documents to Markdown.
Lightweight coding agent that runs in your terminal
A powerful MCP toolkit for coding, providing semantic retrieval and editing capabilities - the IDE for your agent
🤖 AI-powered software engineering multi-agent system with researcher and developer agents that automate code implementation through intelligent planning and execution. Built with LangGraph multi-ag…
An autonomous agent that conducts deep research on any data using any LLM providers
An index of the LangChain + LangGraph ecosystem: concepts, projects, tools, templates, and guides for LLM & multi-agent apps.
Achieve state of the art inference performance with modern accelerators on Kubernetes
[CVPR 2025] MMAudio: Taming Multimodal Joint Training for High-Quality Video-to-Audio Synthesis
Open-source observability for your GenAI or LLM application, based on OpenTelemetry
A PyTorch library for all things Reinforcement Learning (RL) for Combinatorial Optimization (CO)
Recent research papers about Foundation Models for Combinatorial Optimization
[NAACL 2025] AgentMove: A Large Language Model based Agentic Framework for Zero-shot Next Location Prediction.
A list of awesome academic researches and industrial materials about Large Language Model (LLM) and Artificial Intelligence for IT Operations (AIOps).
Benchmark for automated failure attributions in agentic systems (🏆 ICML 2025 Spotlight)
An open-source, code-first Python toolkit for building, evaluating, and deploying sophisticated AI agents with flexibility and control.
Code for the paper "Efficient Training of Language Models to Fill in the Middle"
Official repository for the paper "COAST: Enhancing the Code Debugging Ability of LLMs through Communicative Agent Based Data Synthesis".
A multi-lingual program repair benchmark set based on the Quixey Challenge
The repository for paper "DebugBench: "Evaluating Debugging Capability of Large Language Models".
CodeGeeX: An Open Multilingual Code Generation Model (KDD 2023)