全部标签
目录标签

#information-extraction

公开记录中带有此标签的全部仓库——标签来自其 GitHub 主题或软件包注册表发布的关键词。健康度量遵循与记录其余部分相同的版本化方法论。

9 条记录
标签为“information-extraction”按健康指数排序
PyPI
89优秀健康指数
run-llama/llama-parse-py
Python SDK for OCR and document parsing in the cloud with LlamaParse
Python★ 57↓ 28.6M/月2026年8月8日
MIT2026年8月8日 · 指标 2.10.0
npm
88优秀健康指数
run-llama/llama-parse-ts
Typescript SDK for OCR and document parsing in the cloud with LlamaParse
TypeScript★ 29↓ 248.7K/月2026年8月1日
MIT2026年8月1日 · 指标 2.10.0
PyPI
87优秀健康指数
urchade/GLiNER
Generalist and Lightweight Model for Named Entity Recognition (Extract any entity types from texts)
Python★ 3,531↓ 750.4K/月2026年8月13日
Apache-2.02026年8月13日 · 指标 2.10.0
PyPI
84优秀健康指数
yifanfeng97/Hyper-Extract
Hypergraph is more powerful. Transform unstructured text into structured knowledge with LLMs. Graphs, hypergraphs, and spatio-temporal extractions — with one command.
Python★ 3,223↓ 2,255/月2026年8月1日
自定义许可证2026年8月1日 · 指标 2.10.0
PyPI
78良好健康指数
adbar/htmldate
Fast and robust date extraction from web pages, with Python or on the command-line
Python★ 155↓ 15M/月2026年8月27日
Apache-2.02026年8月27日 · 指标 2.10.0
crates.io
59中等健康指数
fbilhaut/gline-rs
Inference engine for GLiNER models, in Rust
Rust★ 158↓ 8,548/月2026年8月6日
Apache-2.02026年8月6日 · 指标 2.10.0
PyPI
45薄弱健康指数
kpwhri/konsepy
Framework for build NLP information extraction systems using regular expressions.
Python★ 1↓ 194/月2026年9月5日
无许可证2026年9月5日 · 指标 2.10.0
PyPI
31存在风险健康指数
lum-ai/odinson
Odinson is a powerful and highly optimized open-source framework for rule-based information extraction. Odinson couples a simple, yet powerful pattern language that can operate over multiple representations of text, with a runtime system that operates in near real time.
Scala★ 742026年9月5日
Apache-2.02026年9月5日 · 指标 2.10.0
Maven
14危急健康指数
BMDSoftware/neji
Flexible and powerful platform for biomedical information extraction from text
JavaScript · Java★ 392026年8月1日
无许可证2026年8月1日 · 指标 2.10.0