RSS 191 projects tagged "Windows"

Download Website Updated 04 Mar 2014 Sanzang

Screenshot
Pop 312.65
Vit 10.36

Sanzang is a compact and simple cross-platform machine translation system. It is especially useful for translating from the CJK languages (Chinese, Japanese, and Korean), and it is very suitable for working with ancient and otherwise difficult texts. Unlike most other machine translation systems, Sanzang is small and approachable. Any user can develop his or her own translation rules, and these rules are simply stored in a text file and applied at runtime.

Download Website Updated 05 Jun 2011 Pair

Screenshot
Pop 22.74
Vit 35.87

Pair is a program that reads the strings from an input file, pairs them with the strings from a second file, and writes the results to an output file. It doesn't support Unicode, and the comparison function is very simple.

Download No website Updated 05 Sep 2013 Multibyte Keyword Generator

Screenshot
Pop 33.65
Vit 1.44

Multi-byte Keyword Generator extracts meta keywords from multi-byte text. It is an enhanced version of the "Automatic Keyword Generator" class originally written by Ver Pangonilo. This version provides better word segmentation, the ability to handle multi-byte strings, and support for text in multiple languages.

No download Website Updated 12 Apr 2013 Lucidor

Screenshot
Pop 83.16
Vit 5.26

Lucidor is a program for reading and handling e-books. It supports e-books in the EPUB file format and catalogs in the OPDS format.

Download Website Updated 29 Mar 2010 jSmaTeP

Screenshot
Pop 55.48
Vit 2.17

jSmaTeP assists in the use of Java for processing import and export data by configuring a data structure rather than by programming it. The structure of the import data is specified in an XML file. jSmaTeP then generates a value object representing exactly one row or record in the import file based on a given XML data configuration. This means that if the import or export format changes, only the XML data configuration needs to be changed to match it.

Download Website Updated 10 Aug 2010 xMarkup

Screenshot
Pop 25.50
Vit 1.14

xMarkup is a command line and GUI utility for multipurpose processing of a set of text files. It can be used to generate or edit the navigational cross-references within a set of HTML documents, analyze and convert the structure or content of SGML, XML, HTML, or text documents, split or merge text files with specified rules, analyze and extract data, generate scripts, and more. xMarkup supports a built-in procedural language which may be used to describe rules of the processing. This language is a simple dialect of the Icon programming language.

No download Website Updated 26 Apr 2010 Okapi Framework

Screenshot
Pop 34.00
Vit 1.46

The Okapi project’s main purpose is to architect a set of building blocks for the creation of larger open source localization and translation tools. But many Okapi components are generic enough to be of interest to the text mining, natural language processing, and text retrieval communities. Okapi’s many text filters (HTML, Properties, XML (ITS XPath-based rules), OpenXML, ODF, Regex etc.) provide a straightforward way to access the text of multiple document formats. Its document events and pipeline can be made to integrate with other frameworks such as UIMA, LingPipe, OpenPipeline, OpenNLP, GATE, and Lucene. The advantage of Okapi’s text filters is that not only is text extracted, but all non-textual formatting is preserved. It is possible to decompose a document into events, process them via the pipeline, and then rebuild the input document without loss. Structural information can be added to Okapi document events so that tables, lists, links, titles etc. are grouped together and treated as a unit. This is useful when context based on a “universal” document structure is needed. The Okapi event model supports user configurable annotations, similar to UIMA, but simpler and more restricted in scope. User can annotate spans of text or add new resources such as translation memory matches, terminology, token types, or part of speech information.

Download No website Updated 24 Jul 2013 Ascii Design

Screenshot
Pop 93.61
Vit 5.25

Ascii Design is an ASCII art program based on the FIGlet engine. You can create text-based art for many types of decorations for Web sites, email, text files, etc.

Download Website Updated 27 Mar 2009 TextEditor++

Screenshot
Pop 18.00
Vit 1.00

TextEditor++ is a cross-platform text editor for both plain and formatted text files, and for printing. It includes a tool for PDF conversion of plain text files.

No download Website Updated 21 Sep 2012 SILVERCODERS DocStorage

Screenshot
Pop 35.44
Vit 4.06

SILVERCODERS DocStorage is a utility to improve document management. You can have one database for all invoices, guarantees, protocols, and other documents. DocStorage can extract plain text from documents in doc, XLS, PPT, PDF, RTF, ODT, ODS, ODP, docx, XLSX, PPTX, and many other formats. It can use an OCR engine to extract plain text even from scanned documents. It can perform global fulltext search in all documents regardless of format. It supports document versioning, document duplicate detection, document notes, and document signing. It provides full integration with software suites like Microsoft Office and OpenOffice.

Screenshot

Project Spotlight

Synth

A powerful C++ templating framework.

Screenshot

Project Spotlight

pgCluu

A PostgreSQL performance monitoring and auditing tool.