nlp-in-practice
@kavgan
Starter code to solve real world text data problems. Includes: Gensim Word2Vec, phrase embeddings, Text Classification with Logistic Regression, word count with pyspark, simple text preprocessing, pre-trained embeddings and more.
Pijlers en weging
Waarom deze score
- Laatste commit 2125 dagen geleden
- Top 73% naar sterren in AI & ML › Machine-learning-frameworks
Feiten
- Licentie
- Onbekend
- Taal
- Jupyter Notebook
- GitHub-sterren
- 1.186
- Laatste commit
- 2020-12-02 (6 jaaren geleden)
- Laatste release
- Onbekend
- OpenSSF Scorecard
- Nog niet gemeten
- Zelf hosten
- Niet vastgesteld
- Maintainer-locatie
- Onbekend
- Bron
- github-crawl
Alternatieven in Natuurlijke taalverwerking
State-of-the-Art Embeddings, Retrieval, and Reranking
Convert documents to structured data effortlessly. Unstructured is open-source ETL solution for transforming complex documents into clean, structured formats for language models. Visit our website to learn more about our enterprise grade Platform product for production grade workflows, partitioning, enrichments, chunking and embedding.
Proofread more than 20 languages. It finds many errors that a simple spell checker cannot detect.
微舆:人人可用的多Agent舆情分析助手,打破信息茧房,还原舆情原貌,预测未来走向,辅助决策!从0实现,不依赖任何框架。