Navi-Bench: benchmarking web agents on everyday tasks directly on real websites
☆21Sep 14, 2026Updated this week
Alternatives and similar repositories for navi-bench
Users that are interested in navi-bench are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆19Mar 7, 2026Updated 6 months ago
- ☆21Jul 17, 2026Updated 2 months ago
- Lightweight method based on shortest path on word graphs and NLP to generate single sentence summaries that highly relevant and grammatic…☆19Jan 29, 2017Updated 9 years ago
- OmniByteFormer is a generalized Transformer model that can process any type of data by converting it into byte sequences, bypassing tradi…☆17Updated this week
- ☆31Oct 7, 2024Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Code for paper OpenWebRL: Online Multi-Turn Reinforcement Learning for Visual Web Agents☆51Aug 17, 2026Updated last month
- An adaptive training algorithm for residual network☆17Aug 22, 2020Updated 6 years ago
- Interferobot: aligning an optical interferometer by a reinforcement learning agent☆12Nov 22, 2022Updated 3 years ago
- Implementation of a differetiable discrete-time algebraic Riccati equation (DARE) solver in PyTorch.☆11Nov 16, 2022Updated 3 years ago
- A dataset containing features extracted from measurements collected from multi-hop WSN deployments☆11Nov 9, 2016Updated 9 years ago
- A PyTorch Implementation of the Importance Weighted Autoencoders☆39Dec 2, 2018Updated 7 years ago
- [ICLR 2026] BARREL: Boundary-Aware Reasoning for Factual and Reliable LRMs☆18May 21, 2025Updated last year
- This project includes code for using the AsyncWebRL and WebGym frameworks to train web agent models.☆49Jun 9, 2026Updated 3 months ago
- Implementaion of Gaussian Process Recurrent Neural Networks developed in "Neural Dynamics Discovery via Gaussian Process Recurrent Neura…☆40Dec 8, 2022Updated 3 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Implementation of the MIWAE method for deep generative modelling of incomplete data sets.☆41Mar 16, 2024Updated 2 years ago
- Repo reproducing experimental results in "Addressing the Topological Defects of Disentanglement"☆24Jul 15, 2022Updated 4 years ago
- a very fast parser for sparse matrix at libsvm format☆10Nov 13, 2017Updated 8 years ago
- [Code] Deep Multi-task Representation Learning: A Tensor Factorisation Approach☆57Jul 4, 2017Updated 9 years ago
- Implementation of PCA algorithm using Gram-Scmidt modification on NIPALS☆10Jun 13, 2015Updated 11 years ago
- Official Implementation for the paper: A Variational Framework for Improving Naturalness in Generative Spoken Language Models☆24Jun 18, 2025Updated last year
- ☆23Feb 18, 2019Updated 7 years ago
- Tinkering RL☆29Updated this week
- Implementing FastSent in theano☆12May 2, 2016Updated 10 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Jax implementation of VIT-VQGAN☆10Jan 25, 2024Updated 2 years ago
- The Tweets2013 Internet Archive collection☆10Aug 7, 2020Updated 6 years ago
- Text-to-Speech Benchmark☆30Aug 18, 2026Updated last month
- Counterfactual Evaluation and Learning for Interactive Systems: Foundations, Implementations, and Recent Advances☆12Aug 14, 2022Updated 4 years ago
- A context encoder for audio inpainting☆26Mar 24, 2023Updated 3 years ago
- Factoried Personalized Markov Chains for Next Basket Recommendation in R and Python☆14Jan 7, 2018Updated 8 years ago
- ☆10Jul 5, 2016Updated 10 years ago
- ☆14Aug 26, 2016Updated 10 years ago
- Orchard: An Open-Source Agentic Modeling Framework☆522Jul 31, 2026Updated last month
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Efficient implementation of Generative Stochastic Networks☆12Nov 28, 2013Updated 12 years ago
- Source code repo for paper "TLDR: Token Loss Dynamic Reweighting for Reducing Repetitive Utterance Generation"☆10Aug 11, 2023Updated 3 years ago
- Deep Critical Learning. Implementation of ProSelfLC, IMAE, DM, etc.☆31Dec 23, 2022Updated 3 years ago
- Official repository of "Efficient and Effective Query Expansion for Web Search", Short Paper @ CIKM 2018☆15Nov 17, 2019Updated 6 years ago
- Python implementation of CLEAR multi object tracking (MOT) evaluation metrics☆21Jun 24, 2022Updated 4 years ago
- A new paper list for multi-agent reinforcement learning (actively maintained)☆24Mar 27, 2020Updated 6 years ago
- Pairwise Interaction Tensor Factorization☆10Oct 11, 2018Updated 7 years ago