这是一个工具程序集合,方便我们平时对数据进行预处理。针对文本处理的内容较多。包括分词(集成了张华平分词、结巴分词)、文件处理增强(如读取文本到Map中,保存文本到Map)和语料模型(把文档转换成矩阵,就算单词数量等)
☆21Oct 3, 2024Updated last year
Alternatives and similar repositories for HFUTUtils
Users that are interested in HFUTUtils are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 以知乎日报为数据源,全流程实践一个机器学习过程,从数据获取到数据分析,对知乎日报进行聚类、分类,并可视化这一过程☆17Apr 6, 2016Updated 10 years ago
- java编写的文本分词后利用tfidf计算每个文档的单词的tfidf值,并保存到文件中☆17Nov 14, 2017Updated 8 years ago
- 操作系统期末复习用的笔记☆12Feb 27, 2022Updated 4 years ago
- 这是针对大数据集优化了的双数组字典树,使得在大数据集上构建速度也比较满意,查询速度不随数据集的增加而增加,同时解决了数据集需要有序的要求.☆10Feb 7, 2018Updated 8 years ago
- Java 23 种设计模式,包含 Demo 实例,便于理解☆10Feb 15, 2019Updated 7 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- 实现中文文本分类,支持文件、文本分类,基于多项式分布的朴素贝叶斯分类器。由于工作实际应用是二分类,加之考虑到每个分类属性都建立map存储词语向量可能引起的内存问题,所以目前只支持二分类。当然,直接复用这个结构扩展到多分类也是很容易。之所以自己写,主要原因是没有仔细研读mah…☆23Sep 13, 2016Updated 9 years ago
- News classification & recommendation in Keras☆13Jun 15, 2020Updated 5 years ago
- CCL2024 Chinese Essay Rhetoric Recognition and Understanding☆17Oct 1, 2024Updated last year
- 经验构件库(Java版)☆16Mar 7, 2022Updated 4 years ago
- http://disqus.com/api/ API bindings and CLI for NodeJS☆13Jul 5, 2016Updated 9 years ago
- 周末为广大社会青年推荐同城活动、美食、周边游、运动、话剧、音乐会、读书会、展览、户外、酒吧、购物、电影等活动,提供信息,并且能够帮助客户实时在线预订、信息查询、网上支付以及地图定位与导航。 周末为广大商家提供宣传推广,帮助商家增加周末客流量,完善广大商家网上营销的空缺,增加…☆11Nov 21, 2018Updated 7 years ago
- Word2VecfJava: Java implementation of Dependency-Based Word Embeddings and extensions☆17Mar 5, 2018Updated 8 years ago
- 分析优酷、土豆等视频网站的播放页面,获取视频标题、视频截图、M3U8地址以及插入页面的swf地址。☆11May 5, 2014Updated 12 years ago
- 使用 rails + mongodb 搭建论坛☆10Oct 25, 2019Updated 6 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- 搜索引擎原理实践, java, 多线程异步下载,CompletableFuture,下载,分析,倒排索引。☆16Apr 4, 2019Updated 7 years ago
- 兼容pc、wap、原生小程序、Taro框架等的基于canvas炫酷的流式渐变(canvas flow gradient for PC、WAP、minProgram)☆13Jan 7, 2023Updated 3 years ago
- ☆14Apr 12, 2022Updated 4 years ago
- Batch processor to enable large content be digested by Ollama, focused around book processing and translations by default, fully, configu…☆36Oct 27, 2025Updated 6 months ago
- Explore the potential of recommendation system using reinforcement learning☆15Apr 23, 2020Updated 6 years ago
- analyzer adapter for solr 5, we support Jieba, and stranford in the future☆62Sep 3, 2018Updated 7 years ago
- Leap Motion + Oculus Rift + Blender + Python 3 + Arch Linux☆11May 14, 2015Updated 11 years ago
- ☆11Jun 6, 2021Updated 4 years ago
- A general-purpose Java library for performing structured learning.☆23Jul 5, 2022Updated 3 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ☆11May 2, 2017Updated 9 years ago
- Semantic Dependency Parsing Toolkit☆22Jun 26, 2015Updated 10 years ago
- WTP - 一个动态线程池管理系统☆14Aug 16, 2022Updated 3 years ago
- android 自动接听电话和挂断(支持所有版本)☆14Apr 3, 2015Updated 11 years ago
- Bandit algorithms for online learning to rank☆17May 26, 2019Updated 7 years ago
- Text2Neo4j 是一个遍历文档、从文本中提取关系并将其保存到 Neo4j 数据库中以形成知识图谱的工具。本项目结合了 Dify 和 LLaMA3.1(8B 模型)来高效处理和提取复杂关系。☆24Aug 31, 2024Updated last year
- ☆24Mar 3, 2026Updated 2 months ago
- Software for the experiments reported in the RecSys 2019 paper "A Simple Multi-Armed Nearest-Neighbor Bandit for Interactive Recommendati…☆21Apr 4, 2024Updated 2 years ago
- this is a java gateway project,And the project is dependent on spring cloud gateway and spring security.add dynamic route rule support an…☆14May 15, 2019Updated 7 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- General Vectorization Lib for Machine Learning Tools☆31Jul 22, 2016Updated 9 years ago
- 开挂人生重开模拟器(500岁以上成仙的概率大幅提升)☆16Nov 23, 2021Updated 4 years ago
- ☆10Mar 3, 2020Updated 6 years ago
- Hackintosh 黑苹果 macOS Monterey for 微星(MSI)MPG Z490M GAMING EDGE WiFi刀锋板 + intel i7-10700K☆10May 10, 2024Updated 2 years ago
- 程序猿的简单记账!GolenCat---easy to charge!☆14Jun 21, 2018Updated 7 years ago
- Recommendation System using Deep Q-Networks and Double Deep Q-Networks☆13May 23, 2020Updated 6 years ago
- Spark Mllib 1.6.0版本算法封装☆11Mar 8, 2017Updated 9 years ago