What is DirectInfo? It is a web-based tool for searching text in documents. Having DirectInfo, ORACLE and an internet connection, you can easily index and search all the textual resources in your local network or the internet. To learn more about the DirectInfo features just click the topics and find out how to use them. You will find useful information to help you work with DirectInfo.
You can find information in this help system using the text searching mechanism integrated in your browser. After you are inside the help page, you can search for specific text on that page by clicking the Edit menu, and then clicking Find (on this page). A Find dialog will appear and you must just type the text and press Enter. The same dialog could be invoked by pressing Ctrl+F simultaneously.
| This section describes the following topics: how
the search is performed, how to specify advanced
search expressions, how to search for special
characters and keywords and how to use expressions
related to thesauri.. |
|
To enter a query into DirectInfo Documents, just type in a few descriptive words and hit the 'enter' key (or click on the Search button) for a list of relevant hits. When a search is initiated the system searches the contents of the items in the selected Document group(s) and finds those which match the search expression. Then a special score is calculated which represents the relevance of the found items to the search expression. The best-scored items are presented first in the results page. In order to perform an advanced search, please use the following keywords (expressions). |
|
Search expressions can be combined to achieve greater
accuracy
in searching. Following reserved words and special charachters can be
used:
|
Following special charachters can be used as wildcards in
expressions:
|
Following advanced expressions are related to thesauri:
|
| Searching for documents on a particular topic is as easy as typing a question, or just a word or two, into the "Search Options" panel. You can perform simple search or advanced search by defining a set of options. |
|
You can simply type a word and search immediatelly for all the
documents that contain it. Nothing else is needed to be specified. For example, if you want to find information about "test":
Process might become even simple when you repeat recent search or want to search for word listed in Search results page. "Quick keywords typing" describes how Direct Info Documents can help you. |
| You can set the Search Options panel to Advanced
mode,
which lets you narrow down a search by defining a more complex
search query. You can
choose one or more different. document groups; specify some file
options: type, name, size and date; define values of specific
parameters that affect the presentation of the results. For example, click the link Advanced and you will switch to "Advanced Search" mode. Then you could build a more complex query like this:
If you want to limit searching in a particular group or groups of documents follow these steps.
If you want to reset all the fields, check boxes and lists to their initial state, please click the "Reset" button. |
Searching in a limited set of files is possible by using some
file
filtering options. To define a particular file name or extension, file
type, file size, or file date range you must take the following steps:
Note: File options could be defined and used even without search query. This allows searching for specific files that are indexed and could be found by the search engine. |
| In addition to the filtering options, there is a possibility
to define
how the hits (documents that match your search query) will be
displayed in the results page. You can change the presentation options this way:
|
Direct Info Documents assists you typing keywords in search
filed in two ways:
|
| After you press the "Search"
button you will see the
results page. It contains the hits (the documents that match the
search query). The most important elements on the results page are:
This layout of the search result is called list view. It is looks more like the search results page of the most popular internet search engines. |
| Search results could also be represented in a table view with
coulmns showing the properties of each document. This facilitates
sorting the results by a particular column. Each row of the table view consist of:
Note: Document attributes and hits contexts are shown as columns only if they are switched on in the preferences. Each preference that is not shown as column could be seen as hidden attributes while mouse cursor is over the document name. |
Direct Info Documents can index e-mail content including
e-mail attachments in the same way as it process other documents like
files or URLs for example. However it is distinguished not only in the
way it visualizes document’s content, but also in processing additional
information like: From, To, Cc, Bcc, Subject and Attachment name. What
is more, you can ask Direct Info to search these fields using the
reserved word ‘within’ followed by From, To, Cc, Bcc, Subject, Sent,
Attachment, as it is shown in the example below:
Direct Info treats attachments as separate documents, but it keeps dependencies between e-mail and attached file. There are two viewers, which are used for displaying e-mail content: one for e-mail body and one for HTML version of its attachments. Some attachments like pictures or movies are not indexed (which is set in Direct Info configurations page) and their HTML version is not available.
|
| The user must enter either a full path to the document or
part of the document name into the 'Document name' field and then must
click the "Check Status" button. Then, DirectInfo will search for
documents that match the criteria and will report the results in
the right frame. The report is ina tabular format and displays a list
of the full paths to the document, together with the document status.
The document status field shows 'Y' if the docment is indexed or 'N',
if the document is not indexed due to some reason. By clicking the link
with the document path, the user could see the entire status
information about the selected document (the status with explanation of
the error, if any; the document size, date and index time; the name of
the document group that the document comes from. The user could see only the non-indexed documents by chechking the check box "Only not indexed file". NOTE: The file path must be specified in the exact same way as DirectInfo has located it during the crawling process. For example, if the client computer's name is MyPC, and if it has a shared folder
called Docs on the C: drive, then the client can access
the
folder either
directly via C:\Docs\somefile.doc, or through the LAN via
\\MyPC\Docs\somefile.doc. These two paths will
produce the same file.
However, if DirectInfo has located the file through the Initial Path \\MyPC\Docs\,
then the only path that will
make DirectInfo recognize
the file will be the latter one, i.e. \\MyPC\Docs\somefile.doc. |
| Further below, for the user's convenience, a tree of the
Document
Groups and their
respective Initial Paths is displayed. The user can browse what Initial
Paths are available by clicking on the "plus" icon before the name of
the document group. This will cause expanding of this document group
and the Initial Paths could be seen. The user may click on the link of
a specific document group, or some of its initial paths. Then a
list
of all documents in this group will appear in the right frame, together
with their statuses. |
| In the context of DirectInfo
document groups are logicaly defined groups that contain documents from
different kind of sources: file shares, web servers, mail servers. Each
separate source of documents has its own initial path depending on its
type. To manage the document groups click the "Document Groups" item in the main menu and the "Document Groups" page will open. You must have administrative rights to access this page. This screen shows the list of all document groups separated on pages. For each document group, the following is shown:
If a particular process is running for this group (crawling, indexing, optimizing), the group's name is grayed out and the group is disabled for further actions. Content of this page is refreshed every 10 sec. due to "Automatic Refresh" check box (on its right upper side). It might be unchecked, but then you need to refresh page content manually. |
| The indexing process includes tree stages based on the tasks
performed with the documents: crawling, indexing and optimizing. This
indexing process is executed as a database job so that controlling of
it means startting, stopping and resuming of the database job at
these different stages. Here is a short description of the stages: 1. Crawling: At this stage the job will run the respective crawlers depending on the type of the initial paths and they will crawl the files. Crawling could be virtually stopped or paused, but it will perform its entire work because crawlers are external modules and their execution could not be controlled. Stopping or pausing during this stage actually will stop the job at indexing stage. Restarting at this stage will start the crawler from the begining. Resuming will continue with indexing. 2. Indexing: At this stage the job will run the indexing of the entire document group on chunks. The term chunk here means a small set of documents that are processed in a group. After a chunk is indexed the indexing process will index next chunk untill the entire list of documents become indexed. Using this specific technique allows the user to stop, pause and resume indexing. Actually if the user have stopped or paused the indexing, the process will continue to index until the last document in the chunk is processed. Resuming the job will continue to index the next chunk. Restarting the job will start the crawling. 3. Optimizing: At this stage the job will run optimization of the already created indexes. Optimization is actually executed as separate stage and could not be stopped, paused or resumed. Note: The indexing process (job) could be scheduled for execution at a specific time range so that it is a frequent case that some of the jobs will be controlled automaticaly by the schedule controler. See section Schedules for more details. |
You can create a new document group by following these steps:
|
You can change the name and the description of an existing
document group by following these steps:
|
| The initial paths of the document groups are needed
to
instruct the crawler -- the module which scans (or "crawls") a given
location and detects the documents contained therein -- where is the
starting point for the crawling process.
At least one initial path must be specified for each document group in
order for the group to be usable. You can manage the initial paths
for a selected document group after clicking on the group's name in the
"Document groups" page. The list of initial paths for this document group will appear. Each record of the list contains:
On the bottom of the page are located some links related to the current document group. Clicking on these links will display a page for managing the access rights or showing the index logs and statistics for this document group. |
| You can create the following types of initial paths: |
You can create a new URL initial path by following these
steps:
|
You can create a new FILE initial path by following these
steps:
|
You can create a new EMAIL initial path by following these
steps:
|
You can create a new LDAP initial path by following these
steps:
|
You can change an existing initial path by following these
steps:
If you want to reset all fields to their initial values, click the "Reset" button. |
There are 2 different access rights for a document group:
|
| Index logs for document groups store information for the
errors that have occurred during the indexing. Viewing these logs is
useful, if you want to identify the problems that concern certain
documents, which are not indexed. Each index log has its unique start and end date that define the time frame of an indexing session. You can browse the full log or the summary log for an indexing session. |
| Index log summary gives a list of the different types of
errors that have occurred during the indexing session. Next to each
error is the number of the documents that are affected by this error. If you want to see the list of files for a selected error - just click on the link with the error message. |
Full index log shows more detailed information about the
errors that have occurred during the indexing session. Each row from
the
list of errors contains:
|
| You can see the statistics gathered during the indexing
session for each document group. The statistical data is useful for
finding the document types or single documents that consume most of the
indexing time. 2 separate statistical reports are availabale for your
consideration: statistics summary and statistics details. |
Statistics summary report includes the total time for
indexing the entire document group and a list of all indexed document
types. Each row of the list contains:
If you want to see the "Statistics details", just click the link "Details" on the top of the page. |
Statistics details report includes
the total time for indexing the entire document group and a list of all
indexed documents. Each row of the list contains:
If you want to see the "Statistics summary", just click the link "Summary" on the top of the page. |
| You can use DirectInfo to find duplicated files in the file
system (files with same content). Just type some file paths and press the button "Find Duplicates". Then a list of all duplicated files will appear and below each file are listed its duplicates. If you want to search for a specific file name, just type it in the field "Duplicates for". Wild cards are also applicable here. In this page you can see also the duplicates count per document group. These counts are calculated during the indexing process if the duplicates functionality is switched on. |
| There is a large list of user preferences that you could use to adjust DirectInfo to your preferred style. In the left frame you will see the groups of preferences combined by similarity. Clicking on the name of each group will show in the right frame a list with the preferences of this group. |
Each row of the preferences list includes:
|
Here is a short description of each of the user preferences:
|
You can change a preference's value by doing the following:
|
| Application parameters are used to define some values used
by the application and to control the application behaviour in certain
situations. This page is accessible only if you have administrative
rights. In the left frame you will see the groups of application
configuration parameters combined by similarity. Clicking on the name
of each group will show in the right frame a list with the application
parameters of this group. |
Each row of the parameters list includes:
|
Here is a short description of each of the application
parameters:
|
You can change a parameter's value by doing the following:
If you want to reset all fields to their initial values, click the "Reset" button. |
| The schedules are useful, if, for example, you want to crawl
and
reindex a document group in a certain time frame and repeat this
process regularly, or if you want to specify a time when DirectInfo
should optimize itself by gathering statistics of all its data. This
page is accessible only if you have administrative rights. |
To create a new schedule follow these steps:
If you want to clear all fields, click the "Clear" button. Here are the different forms for the different types of schedules: |
To change an existing schedule follow these steps:
|
| Shedule controller is used to manage all scheduling processes. When it is run, it monitors defined shedules and runs them if it is necessary. So when it is stopped no one task will run. Use this menu to stop or start controller. |
| To restrict the user rights for certain actions and document
groups, the administrator must implement a security policy. The "Security administration" page can
be reached from the main menu. It is accessible only if you have
administrative privileges. It includes two main sections: |
Managing the users includes creating, editing and deleting a
user, assigning a
user to a user group, managing user access rights for a document group.
All
these actions can be started from the "Managing
users" page. It shows a
list of all users separated on pages. Each row contains:
|
You can create a new user by performing the following steps:
NOTE: When the application works with Basic authentication schema, the user created, should also exists as Oracle user. |
You can change an existing user by performing the following
steps:
|
You can edit the groups that an existing user is a member of.
To do this, follow these steps:
If you want to exclude the user from a user group, do these steps:
Note that the movement of a user group is immediate and does not need to be confirmed! |
Each user can be given access rights to a specific document
group. You can do this by following these steps:
|
You can attach an existing permission set to user in order to
enable
its permission statements to be applied to this user. Just follow
these steps:
|
You can detach an existing permission set from user in order
to disable
its permission statements to be applied to this user. Just follow
these steps:
|
Mapping network users to DirectInfo users is important when
you have shared network files and you need to apply the security policy
of the operating system (OS) concerning these files. Supposing that you
already have the users imported from the OS to DirectInfo, it is
important to allow these users to login to DirectInfo and use its
features. That is why you need to map these users to DirectInfo users.
To do this just follow these steps:
|
Managing the users groups includes creating and deleting a
user
group, assigning
users to a user group, assigning roles to user groups, managing user
group access rights for a document group. All these actions can be
started from the "User Groups"
page. It shows a list of all users separated on pages. Each row
contains:
|
You can create a new user group by performing the following
steps:
If you want to clear all fields, click the "Clear" button. |
You can edit the users of an existing user group. To do this
follow these steps:
If you want to exclude user from the user group do these steps:
Note that the movement of a user is immediate and does not need to be confirmed! |
You can edit the roles of an existing user group. To do this
follow these steps:
If you want to exclude a role from the user group do these steps:
Note that the movement of a role is immediate and does not need to be confirmed! |
Each user group can be granted access rights to a specific
document group. You can do this by following these steps:
|
| This page shows a summary of all usage statistics, i.e. which
objects have been used, how many times, the min/max/average elapsed
times,
etc. By clicking on an object, the user will be taken to the Object Usage Details page for that
object and will be shown further details regarding the usage of that
object. It is also possible to select a time period for which you want to see usage statistics. This can be done by selecting a 'Start date' and an 'End date' and clicking on the 'Search' button. |
| This page shows a summary of all instances when a certain
object
has been used. The information includes the start and end times as well
as
the time it took for the object to execute. The page also shows which
user
used the object, whether the execution completed successfully, and some
additional details that may have been reported by the object when it
was
executed. By clicking on the start time, the user will be taken to the Usage Session Details page where information about the concrete usage of the object will be shown. Sometimes the usage of an object involves the usage of other objects. For example, when searching for given keywords, the search results will usually contain fragments of each document that contains the keywords. The code that generates these fragments is called by our 'object' (whose usage the user is currently viewing). Therefore, the fragments are included into the same 'usage session' as parent current object. By clicking on the object's start time, you will see all other objects that have been invoked as part of the object's session. It is also possible to select a time period for which you want to see usage statistics. This can be done by selecting a 'Start date' and an 'End date' and clicking on the 'Search' button. |
| This page shows a summary of all objects belonging to a given 'usage session'. The information includes the name of each object, the start and end times as well as the time it took for the object to execute. The page also shows which user used the object, whether the execution completed successfully, and some additional details that may have been reported by the object when it was executed. |
| This page shows a statistics for the most used search
queries. The information includes the keywords and the number of calls
for each keyword. The user could filter the results for a specific time
interval between start and end date and specify the number of queries
to be included in the report. |
| DirectInfo has a specific mechanism for granting visibility
to the documents that are accesible to its users. You could define a
set of access permissions that define which documents could be accessed
and searched by a specific users. This set is called permission set in the termsof DirectInfo. It is vital to understand what realy is the permission set in order to be able to manage the user access to the documents. Each permission set is assigned to a single user (called owner of this permission set) and the other users could be attached to or detached from this permission set. This allows to apply the already existing permission set to a list of users when searching text in documents. The permission statements has its own name and description to facilitate the administration. The permission set is comprised of one or more permission statements. The permission statement defines the visibility for specific set of documents called content collection in the terms of DirectInfo. Each content collection has its own name (custom or automatically generated) and type. The available types are:
|
Managing the permission sets includes creating, editing and
deleting a permission set, attaching users to it or detaching users
from it. All these actions can be triggered from the "Permission Sets"
page, reachable from the main menu "Security/Permission Sets". This
page shows a list of permission sets separated on pages. Each row
contains:
|
You can create a new permission set by following
these steps:
If you want to reset all the fields to their initial state, click the "Reset" button. |
You can edit the properties of an existing permission set by
following
these steps:
|
You can delete an existing permission set by following
these steps:
|
You can attach users to an existing permission set to enable
its permission statements to these users. Just follow
these steps:
|
You can detach users from an existing permission set to
disable its permission statements to these users. Just follow
these steps:
|
| Managing the permission statements includes creating and
deleting permission statements. These actions can be triggered from the
"Permission Statements"
page. This
page could be reached only from the page "Permission sets" by clicking on the
name of the selected permission set. The "Permission stataments" page shows a list of permission statements separated on pages. Each row contains:
|
Creating the permission statements is the most important task
while managing permissions to the documents in DirectInfo. To create
new permission statement you must complete the following steps:
The creation of the permission statement could be triggered from different starting points depending on which part of the application you are currently browsing. These starting points have been added to make this process more intuitive and easier for administration. Note that you could perform this operation only if you have a role "ADMIN". To create new permission stament follow these steps:
|
You could create new permission statements directly from the
page with search results. If you have a role "ADMIN" the serach results page is
displayed in a diffrent way than the normal page. There is a check box
next to each document that is shown in the results. You may
select some documents by checking their check boxes. In order to create
permission statement from search results, just do the following:
|
You could create new permission
statements for selected document groups. There is a check box next to
each document group. You may select some document groups by checking
their check boxes. In order to create permission statement from
document groups, just do the following:
|
You could create new permission
statements for selected initial paths of one document group. There is a
check box next to each initial path. You may select some initial paths
by checking
their check boxes. In order to create permission statement from initial
paths, just do the following:
|
You could create new permission
statements for selected users. There is a check box next to each user
name. You may select some users by checking
their check boxes. In order to create permission statement for some
users, just follow these steps:
|
You can edit the name and the description of an existing
permission statement by following
these steps:
If you want to reset all the fields to their initial state, click the "Reset" button. |
You can delete an existing permission statement by following
these steps:
|
Managing the content collections includes adding new items
(documents, document groups, initial paths) to collection and
deleting deleting items from collection. These actions can be triggered
from the "Permission statements"
page by clicking the link "Change
collection". This will open a new page with title "Collection items" showing the items
in the selected colleciton. This page has the following elements:
Note that you cannot add content items different than the type of the collection, because it is not allowed to have mixed type of content collections. |
| In order to change the items in content collection you are
allowed to add new documents, document groups or initial paths to it. To add new collection items you must complete the following steps:
|
You can delete items from content collection by following
these steps:
|