<?xml version="1.0" encoding="UTF-8"?>
<rss xmlns:content="http://purl.org/rss/1.0/modules/content/" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:rdf="http://www.w3.org/1999/02/22-rdf-syntax-ns#" xmlns:taxo="http://purl.org/rss/1.0/modules/taxonomy/" version="2.0">
  <channel>
    <title>topic Re: How the Indexer works in Alfresco Archive</title>
    <link>https://connect.hyland.com/t5/alfresco-archive/how-the-indexer-works/m-p/176871#M130001</link>
    <description>&lt;HTML&gt;&lt;HEAD&gt;&lt;/HEAD&gt;&lt;BODY&gt;&lt;SPAN&gt;Hello,&lt;/SPAN&gt;&lt;BR /&gt;&lt;BR /&gt;&lt;SPAN&gt;If you want to see what's indexed, you may use two ways:&lt;/SPAN&gt;&lt;BR /&gt;&lt;SPAN&gt;- use Luke (Lucene tool which helps you to see you index content)&lt;/SPAN&gt;&lt;BR /&gt;&lt;SPAN&gt;- see the txt files generated in the tomcat temp/Alfresco directory&lt;/SPAN&gt;&lt;BR /&gt;&lt;BR /&gt;&lt;SPAN&gt;Greetings&lt;/SPAN&gt;&lt;/BODY&gt;&lt;/HTML&gt;</description>
    <pubDate>Fri, 26 Jun 2009 08:53:23 GMT</pubDate>
    <dc:creator>nvir</dc:creator>
    <dc:date>2009-06-26T08:53:23Z</dc:date>
    <item>
      <title>How the Indexer works</title>
      <link>https://connect.hyland.com/t5/alfresco-archive/how-the-indexer-works/m-p/176870#M130000</link>
      <description>Hi,i am Running Alfresco-Labs 3 Stable on Debian 4.0 x86_64 with Mysql5 and Tomcat Bundle.Currently i am testing with Abby Finereader and create text below picture pdf's and saving them with the CIFS in Alfresco.Alfresco indexes some word's other's not, on what does that depend?greetingsthomas</description>
      <pubDate>Sat, 31 Jan 2009 22:40:30 GMT</pubDate>
      <guid>https://connect.hyland.com/t5/alfresco-archive/how-the-indexer-works/m-p/176870#M130000</guid>
      <dc:creator>stegbth</dc:creator>
      <dc:date>2009-01-31T22:40:30Z</dc:date>
    </item>
    <item>
      <title>Re: How the Indexer works</title>
      <link>https://connect.hyland.com/t5/alfresco-archive/how-the-indexer-works/m-p/176871#M130001</link>
      <description>&lt;HTML&gt;&lt;HEAD&gt;&lt;/HEAD&gt;&lt;BODY&gt;&lt;SPAN&gt;Hello,&lt;/SPAN&gt;&lt;BR /&gt;&lt;BR /&gt;&lt;SPAN&gt;If you want to see what's indexed, you may use two ways:&lt;/SPAN&gt;&lt;BR /&gt;&lt;SPAN&gt;- use Luke (Lucene tool which helps you to see you index content)&lt;/SPAN&gt;&lt;BR /&gt;&lt;SPAN&gt;- see the txt files generated in the tomcat temp/Alfresco directory&lt;/SPAN&gt;&lt;BR /&gt;&lt;BR /&gt;&lt;SPAN&gt;Greetings&lt;/SPAN&gt;&lt;/BODY&gt;&lt;/HTML&gt;</description>
      <pubDate>Fri, 26 Jun 2009 08:53:23 GMT</pubDate>
      <guid>https://connect.hyland.com/t5/alfresco-archive/how-the-indexer-works/m-p/176871#M130001</guid>
      <dc:creator>nvir</dc:creator>
      <dc:date>2009-06-26T08:53:23Z</dc:date>
    </item>
    <item>
      <title>Re: How the Indexer works</title>
      <link>https://connect.hyland.com/t5/alfresco-archive/how-the-indexer-works/m-p/176872#M130002</link>
      <description>&lt;HTML&gt;&lt;HEAD&gt;&lt;/HEAD&gt;&lt;BODY&gt;&lt;SPAN&gt;We have the same problem and i found out after some tests that there is a problem when adding such a document by CIFS. When you add such an OCRed PDF via CIFS or when you add such a document via the Web-Client or an other client to Alfresco the search results are different.&lt;/SPAN&gt;&lt;BR /&gt;&lt;BR /&gt;&lt;SPAN&gt;The document added by CIFS can not found with the same search string like the same document added via the client - often (maybe allways) it helps when you make a wildcard search and when you e.g. search for&amp;nbsp; &lt;/SPAN&gt;&lt;STRONG&gt;vienna &lt;/STRONG&gt;&lt;SPAN&gt;you have to use &lt;/SPAN&gt;&lt;STRONG&gt;*vienna*&lt;/STRONG&gt;&lt;SPAN&gt; to find the document. Also the search for phrases using e.g.&lt;/SPAN&gt;&lt;STRONG&gt; "vacation in vienna"&lt;/STRONG&gt;&lt;SPAN&gt; does not works with such PDF´s added by CIFS but the same search function works correct if this file was added via the web-client.&lt;/SPAN&gt;&lt;BR /&gt;&lt;BR /&gt;&lt;SPAN&gt;This is a very strange problem which causes a lot of confusion during the tests and use of the system because one of the most important things of such a ECM - "E" for enterprise should be to be able to find the documents you added to the repository - and not for xx% but allways for 100%.&lt;/SPAN&gt;&lt;BR /&gt;&lt;BR /&gt;&lt;SPAN&gt;See also here &lt;/SPAN&gt;&lt;A href="http://forums.alfresco.com/en/viewtopic.php?f=3&amp;amp;t=19701" rel="nofollow noopener noreferrer"&gt;http://forums.alfresco.com/en/viewtopic.php?f=3&amp;amp;t=19701&lt;/A&gt;&lt;SPAN&gt; or &lt;/SPAN&gt;&lt;BR /&gt;&lt;A href="http://forums.alfresco.com/en/viewtopic.php?f=16&amp;amp;t=17306" rel="nofollow noopener noreferrer"&gt;http://forums.alfresco.com/en/viewtopic.php?f=16&amp;amp;t=17306&lt;/A&gt;&lt;SPAN&gt; or&lt;/SPAN&gt;&lt;BR /&gt;&lt;A href="http://forums.alfresco.com/en/viewtopic.php?f=9&amp;amp;t=19341&amp;amp;start=0" rel="nofollow noopener noreferrer"&gt;http://forums.alfresco.com/en/viewtopic.php?f=9&amp;amp;t=19341&amp;amp;start=0&lt;/A&gt;&lt;BR /&gt;&lt;BR /&gt;&lt;SPAN&gt;Where can be the problem&lt;/SPAN&gt;&lt;BR /&gt;&lt;BR /&gt;&lt;SPAN&gt;- problem with language settings we use a german XP to access Alfresco CIFS and add PDF files via drag &amp;amp; drop and we use "english" as selected language for the Web-Client ?&lt;/SPAN&gt;&lt;BR /&gt;&lt;SPAN&gt;- problem with the PDF to text extraction &lt;/SPAN&gt;&lt;BR /&gt;&lt;SPAN&gt;- ……&lt;/SPAN&gt;&lt;BR /&gt;&lt;BR /&gt;&lt;SPAN&gt;Same problem with enterprise 3.0 and Community 3.2 version.&lt;/SPAN&gt;&lt;BR /&gt;&lt;BR /&gt;&lt;SPAN&gt;Any ideas how to solve this ? there seems to be some topics in the forum regarding this problem.&lt;/SPAN&gt;&lt;/BODY&gt;&lt;/HTML&gt;</description>
      <pubDate>Tue, 21 Jul 2009 19:17:21 GMT</pubDate>
      <guid>https://connect.hyland.com/t5/alfresco-archive/how-the-indexer-works/m-p/176872#M130002</guid>
      <dc:creator>wmay</dc:creator>
      <dc:date>2009-07-21T19:17:21Z</dc:date>
    </item>
    <item>
      <title>Re: How the Indexer works</title>
      <link>https://connect.hyland.com/t5/alfresco-archive/how-the-indexer-works/m-p/176873#M130003</link>
      <description>&lt;HTML&gt;&lt;HEAD&gt;&lt;/HEAD&gt;&lt;BODY&gt;&lt;SPAN&gt;The idea is to see what's in the lucene index with Luke (&lt;/SPAN&gt;&lt;A href="http://www.getopt.org/luke/" rel="nofollow noopener noreferrer"&gt;http://www.getopt.org/luke/&lt;/A&gt;&lt;SPAN&gt;) when the document is added through the web interface and then when you add it through CIFS.&lt;/SPAN&gt;&lt;BR /&gt;&lt;BR /&gt;&lt;SPAN&gt;Have also a look in the metadata of the document through the Alfresco node explorer, and checks the field content (which contains something about the language which may be used by the indexer).&lt;/SPAN&gt;&lt;BR /&gt;&lt;BR /&gt;&lt;SPAN&gt;Hope this helps,&lt;/SPAN&gt;&lt;BR /&gt;&lt;SPAN&gt; Alain&lt;/SPAN&gt;&lt;/BODY&gt;&lt;/HTML&gt;</description>
      <pubDate>Tue, 21 Jul 2009 19:35:38 GMT</pubDate>
      <guid>https://connect.hyland.com/t5/alfresco-archive/how-the-indexer-works/m-p/176873#M130003</guid>
      <dc:creator>nvir</dc:creator>
      <dc:date>2009-07-21T19:35:38Z</dc:date>
    </item>
  </channel>
</rss>

