Hi there, DocFetcher doesn't support OST files, and those aren't newer versions of PST files. The commercial big brother of DocFetcher, DocFetcher Pro, does support OST files. However, the website says explicitly that PST files should be preferred for indexing. Prefer PST to OST: While DocFetcher Pro and DocFetcher Server can read OST files to some extent, be warned that OST files are actually just cache files where Outlook temporarily stores some portion of the data from an online account for offline...
No progress feedback while the crawler walks directories during indexing
Hi there, DocFetcher is no longer being developed, it'll only receive bugfixes in the future. Thus, I'm afraid I have to decline the offer. As for the work you've already put in, please note that if you publish your changes, they must be put under the Eclipse Public License. On the other hand, if you keep the changes private, no such requirement applies. Regards q:-) <= Quang
DocFetcher Pro has a different index format, and it's not publicly documented, but it's not encrypted or obfuscated either. It's just a standard Lucene index plus XML files.
You've asked some really complicated questions, so I had Claude Opus 5 look at the source code and write an accurate answer on my behalf: Short answer: the agent can start the queue, and unfortunately that's the least of the problem. The reason your test task sat there without running is that new indexes are added in a "not ready" state — normally it's the indexing dialog that marks them ready when you click through it. But the call that does that is available over the same Py4J connection. An agent...
Hi Mark, The "List Documents" action is non-recursive — it will list only the indexed documents in the selected folder, not in any subfolders beneath that folder. Regards q:-) <= Quang
Illegal Argument Exception
Hi there, The "zip extensions" field is for zip archives only. 7z archives are not zip archives. Regards q:-) <= Quang