π€ Your Intelligent Copilot for Data Exploration and Processing Pipeline
β55Aug 6, 2026Updated this week
Alternatives and similar repositories for data-juicer-agents
Users that are interested in data-juicer-agents are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Data processing for and with foundation models! π π π½ β‘οΈ β‘οΈπΈ πΉ π·β6,847Updated this week
- β28Aug 9, 2025Updated last year
- A collection of ready-to-use Python sample agents built with AgentScope and AgentScope Runtime, covering use cases from CLI tools to fullβ¦β336Apr 10, 2026Updated 4 months ago
- Zero and Few-shot document level relation extraction / β οΈ Development moved to: https://github.com/cea-list-lasti/glidreβ18Mar 13, 2026Updated 4 months ago
- Trinity-RFT is a general-purpose, flexible and scalable framework designed for reinforcement fine-tuning (RFT) of large language models (β¦β682Jul 31, 2026Updated last week
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Structured Pruning Adapters in PyTorchβ19Aug 30, 2023Updated 2 years ago
- β14Nov 22, 2024Updated last year
- β13May 16, 2019Updated 7 years ago
- β14Oct 21, 2024Updated last year
- β10Oct 18, 2021Updated 4 years ago
- The Dataset and Official Implementation for <The ELCo Dataset: Bridging Emoji and Lexical Composition> @ LREC-COLING 2024β16May 11, 2024Updated 2 years ago
- We define and estimate smooth unique information of samples with respect to classifier weights and predictions. We compute these quantitiβ¦β11Mar 9, 2021Updated 5 years ago
- Provide performance insight capabilities for RL frameworks.β53Updated this week
- Top Picks for Data Science Self-Study: From Newbies to Pros!β11Apr 2, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient β’ AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- An implementation of the DISP-LLM method from the NeurIPS 2024 paper: Dimension-Independent Structural Pruning for Large Language Models.β25Aug 6, 2025Updated last year
- PKUηεδΊζη«θͺε¨ε‘«εε¨β21Jan 16, 2021Updated 5 years ago
- Spectral Sphere Optimizerβ132Mar 23, 2026Updated 4 months ago
- Enjoy dark mode on Arxiv papersβ22Aug 16, 2021Updated 4 years ago
- The Dataset and Official Implementation for <Discursive Socratic Questioning: Evaluating the Faithfulness of Language Modelsβ Understandiβ¦β19Aug 7, 2024Updated 2 years ago
- [AAAI 2025] Augmenting Math Word Problems via Iterative Question Composing (https://arxiv.org/abs/2401.09003)β23Oct 2, 2025Updated 10 months ago
- The official implementation of HPCA 2025 paper, Prosperity: Accelerating Spiking Neural Networks via Product Sparsityβ41Aug 9, 2025Updated last year
- A reasoning assistant for your STEM educationβ24Mar 11, 2025Updated last year
- NeuroSync: A Scalable and Accurate Brain Simulation System using Safe and Efficient Speculation (HPCA 2022)β14Nov 9, 2022Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI β’ AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [AAAI 2025] PAT: Pruning-Aware Tuning for Large Language Modelsβ37Feb 1, 2025Updated last year
- Reinforcing General Reasoning without Verifiersβ102Jun 24, 2025Updated last year
- Convert MathML to Latex for OneNote to Markdownβ15Mar 17, 2026Updated 4 months ago
- ICML2019 Accepted Paper. Overcoming Multi-Model Forgettingβ14Jun 5, 2019Updated 7 years ago
- The OlymMATH datasetβ24Jun 1, 2025Updated last year
- [MM 2022] MM-ALT: A Multimodal Automatic Lyric Transcription System (Oral, Top paper award)β21Mar 16, 2024Updated 2 years ago
- The rule-based evaluation subset and code implementation of Omni-MATHβ29Dec 23, 2024Updated last year
- [ACL 2024] Masked Thought: Simply Masking Partial Reasoning Steps Can Improve Mathematical Reasoning Learning of Language Modelsβ27Jul 9, 2024Updated 2 years ago
- Official implementation of Stochastic Taylor Derivative Estimator (STDE) NeurIPS2024β127Nov 27, 2024Updated last year
- Virtual machines for every use case on DigitalOcean β’ AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Official codebase for the ACL 2025 Findings paper: Optimized Text Embedding Models and Benchmarks for Amharic Passage Retrieval.β21Jul 26, 2025Updated last year
- β13Nov 6, 2021Updated 4 years ago
- [ICLR 2025] When Attention Sink Emerges in Language Models: An Empirical View (Spotlight)β165Jul 8, 2025Updated last year
- Pile Deduplication Codeβ18May 15, 2023Updated 3 years ago
- [ACL 2025 main] SCAR: Data Selection via Style Consistency-Aware Response Ranking for Efficient Instruction-Tuning of Large Language Modeβ¦β39Aug 6, 2025Updated last year
- β82Apr 18, 2024Updated 2 years ago
- Python MCP client + server exampleβ26Mar 9, 2025Updated last year