Apache Tika

From COPTR
Revision as of 19:24, 12 November 2013 by MediaWiki default (talk | contribs)
Jump to navigation Jump to search
Detects and extracts metadata and text content from documents.
Homepage:http://tika.apache.org/
License:Apache License, Version 2.0


Description

Java based tool for detecting and extracting metadata and text content from documents.

Searching for Tika on OPF Labs

{search:query=Tikatype=page}

User Experiences

Development Activity

Error in widget Ohloh Project: unable to write file /var/www/html/extensions/Widgets/compiled_templates/wrt6649936e3bcc18_23355061



Release Feed

Link to any RSS feed that is updated when new releases occur, if any, e.g: Failed to load RSS feed from http://projects.apache.org/feeds/rss/tika.xml: There was a problem during the HTTP request: 404 Not Found

Activity Feed

Link to any RSS feed that is updated when issue or code updates occur, if any, e.g:

2024-05-19 05:51:02
Suxing Lee created

When we use `FlinkDynamoDBStreamsConsumer` in `flink-connector-aws/flink-connector-kinesis` to consume dynamodb stream data, there is an out-of-order problem.
Th...

by Suxing Leehttps://issues.apache.org/jira/secure/ViewProfile.jspa?name=Suxing+LeeSuxing Leehttp://activitystrea.ms/schema/1.0/person
2024-05-19 05:48:24
ASF GitHub Bot created a link from ASF GitHub Bot created a link from ASF GitHub Bot updated a link from ASF GitHub Bot changed the Labels to 'pull-request-available'...
by ASF GitHub Bothttps://issues.apache.org/jira/secure/ViewProfile.jspa?name=githubbotgithubbothttp://activitystrea.ms/schema/1.0/person
2024-05-19 05:24:15
ASF GitHub Bot created a link from ASF GitHub Bot updated 2 fields of