Saturday, April 29, 2006
TREC project
Introduction
Trec stand for Text Retrieval Conferences. It is retrieval evaluation experiments.Lancaster mentions that probably the first evaluation study in information retrieval was conducted in1953. More recent evaluation studies have been discussed in the annual review of information science and technology some of the recent retrieval evaluation experiments known as Trec experiment
Text Retrieval Conferences
Researchers in information retrieval have concatenate their research on small collection each of the order of thousand of document .The major problem for the researchers was get a text collection large enough to match the real life situation with an infrastructure adequate for conducting with an conducting test on them .In 1991 in order to order t alleviate the difficulty the US Defiance Advanced Research Projects Agency (DARPA) decided to found the TREC the experiment. National Institution Science and Technology (NIST), in order to Annabel information retrieval research to scale up from small collections of data to large experiments.
Smitten and harman- mention that the goals for the TREE experiment have been to :
1) Increase research in information retrieval on large scale test collections
2) Increase communications among academia, industry and government through on open forum
3) Increase technology transfer between research and products
4) Provide a state of the art showcase of retrieval methods of TREC sponsors
5) Improve evaluation techniques
Over a million documents have been used in TREC I, draw mainly from newspapers, newswires and selected journals. The documents range in size: while most of them range between 300-400 terms, some of them several hundred pages. All documents are uniformly formatted into Standard Generalized Markup Language and distributed CD-ROMs.
TREC has sets of activities
the main activity (core in TREC jargon)
subsidiary activities(tracks in TREC jargon)
the core has two types of tasks
Ad-hoc (that corresponds to retrospective retrieval)
Routing (that corresponds to the selective dissemination of information)
Relevance judgement for the old collections and old topic selection are made available in each subsequent TREC series experiment. The routing task involves using some of the old topics on the new collections: the relevance information from the old collection may be use to help formulate the query or the profile.
Out put list from each participating research team are sent to the NIST. Where they are merged for evaluation for each topic the hundred top ranking documents from all the participating teams are merged into a signal set which is then given to the assessor for relevance evaluation. This method called pooling.
First TREC was stated at 1992. From 1992-2005 it held 14th conferences in November month each and every year except TREC- 2. These conferences are co sponsored by DARPA&NIST.
The 15th TREC conference will started in the month of November in this year
Benefits of TREC
Boolean retrieval
passage or paragraph retrieval
combining the results of more than one search
retrieval based on prior relevance assessment
query expansion & query reduction
String & concept based searching.
Dictionary based stemming. And so on.
Introduction
Trec stand for Text Retrieval Conferences. It is retrieval evaluation experiments.Lancaster mentions that probably the first evaluation study in information retrieval was conducted in1953. More recent evaluation studies have been discussed in the annual review of information science and technology some of the recent retrieval evaluation experiments known as Trec experiment
Text Retrieval Conferences
Researchers in information retrieval have concatenate their research on small collection each of the order of thousand of document .The major problem for the researchers was get a text collection large enough to match the real life situation with an infrastructure adequate for conducting with an conducting test on them .In 1991 in order to order t alleviate the difficulty the US Defiance Advanced Research Projects Agency (DARPA) decided to found the TREC the experiment. National Institution Science and Technology (NIST), in order to Annabel information retrieval research to scale up from small collections of data to large experiments.
Smitten and harman- mention that the goals for the TREE experiment have been to :
1) Increase research in information retrieval on large scale test collections
2) Increase communications among academia, industry and government through on open forum
3) Increase technology transfer between research and products
4) Provide a state of the art showcase of retrieval methods of TREC sponsors
5) Improve evaluation techniques
Over a million documents have been used in TREC I, draw mainly from newspapers, newswires and selected journals. The documents range in size: while most of them range between 300-400 terms, some of them several hundred pages. All documents are uniformly formatted into Standard Generalized Markup Language and distributed CD-ROMs.
TREC has sets of activities
the main activity (core in TREC jargon)
subsidiary activities(tracks in TREC jargon)
the core has two types of tasks
Ad-hoc (that corresponds to retrospective retrieval)
Routing (that corresponds to the selective dissemination of information)
Relevance judgement for the old collections and old topic selection are made available in each subsequent TREC series experiment. The routing task involves using some of the old topics on the new collections: the relevance information from the old collection may be use to help formulate the query or the profile.
Out put list from each participating research team are sent to the NIST. Where they are merged for evaluation for each topic the hundred top ranking documents from all the participating teams are merged into a signal set which is then given to the assessor for relevance evaluation. This method called pooling.
First TREC was stated at 1992. From 1992-2005 it held 14th conferences in November month each and every year except TREC- 2. These conferences are co sponsored by DARPA&NIST.
The 15th TREC conference will started in the month of November in this year
Benefits of TREC
Boolean retrieval
passage or paragraph retrieval
combining the results of more than one search
retrieval based on prior relevance assessment
query expansion & query reduction
String & concept based searching.
Dictionary based stemming. And so on.
Semenar
TREC project
Introduction
Trec stand for Text Retrieval Conferences. It is retrieval evaluation experiments.Lancaster mentions that probably the first evaluation study in information retrieval was conducted in1953. More recent evaluation studies have been discussed in the annual review of information science and technology some of the recent retrieval evaluation experiments known as Trec experiment
Text Retrieval Conferences
Researchers in information retrieval have concatenate their research on small collection each of the order of thousand of document .The major problem for the researchers was get a text collection large enough to match the real life situation with an infrastructure adequate for conducting with an conducting test on them .In 1991 in order to order t alleviate the difficulty the US Defiance Advanced Research Projects Agency (DARPA) decided to found the TREC the experiment. National Institution Science and Technology (NIST), in order to Annabel information retrieval research to scale up from small collections of data to large experiments.
Smitten and Harman- mention that the goals for the TREE experiment have been to :
1) Increase research in information retrieval on large scale test collections
2) Increase communications among academia, industry and government through on open forum
3) Increase technology transfer between research and products
4) Provide a state of the art showcase of retrieval methods of TREC sponsors
5) Improve evaluation techniques
Over a million documents have been used in TREC I, draw mainly from newspapers, newswires and selected journals. The documents range in size: while most of them range between 300-400 terms, some of them several hundred pages. All documents are uniformly formatted into Standard Generalized Markup Language and distributed CD-ROMs.
TREC as two sets of activites
the main activity (core in TREC jargon)
subsidiary activities(tracks in TREC jargon)
the core has two types of tasks
Ad-hoc (that corresponds to retrospective retrieval)
Routing (that corresponds to the selective dissemination of information)
Relevance judgement for the old collections and old topic selection are made available in each subsequent TREC series experiment. The routing task involves using some of the old topics on the new collections: the relevance information from the old collection may be use to help formulate the query or the profile.
Out put list from each participating research team are sent to the NIST. Where they are merged for evaluation for each topic the hundred top ranking documents from all the participating teams are merged into a signal set which is then given to the assessor for relevance evaluation. This method called pooling.
First TREC was stated at 1992. From 1992-2005 it held 14th conferences in November month each and every year except TREC- 2. These conferences are co sponsored by DARPA&NIST.The 15th TREC conference will started in the month of November in this year
Benefits of TREC
Boolean retrieval
passage or paragraph retrieval
combining the results of more than one search
retrieval based on prior relevance assessment
query expansion & query reduction
String & concept based searching.
Dictionary based stemming. And so on.
Introduction
Trec stand for Text Retrieval Conferences. It is retrieval evaluation experiments.Lancaster mentions that probably the first evaluation study in information retrieval was conducted in1953. More recent evaluation studies have been discussed in the annual review of information science and technology some of the recent retrieval evaluation experiments known as Trec experiment
Text Retrieval Conferences
Researchers in information retrieval have concatenate their research on small collection each of the order of thousand of document .The major problem for the researchers was get a text collection large enough to match the real life situation with an infrastructure adequate for conducting with an conducting test on them .In 1991 in order to order t alleviate the difficulty the US Defiance Advanced Research Projects Agency (DARPA) decided to found the TREC the experiment. National Institution Science and Technology (NIST), in order to Annabel information retrieval research to scale up from small collections of data to large experiments.
Smitten and Harman- mention that the goals for the TREE experiment have been to :
1) Increase research in information retrieval on large scale test collections
2) Increase communications among academia, industry and government through on open forum
3) Increase technology transfer between research and products
4) Provide a state of the art showcase of retrieval methods of TREC sponsors
5) Improve evaluation techniques
Over a million documents have been used in TREC I, draw mainly from newspapers, newswires and selected journals. The documents range in size: while most of them range between 300-400 terms, some of them several hundred pages. All documents are uniformly formatted into Standard Generalized Markup Language and distributed CD-ROMs.
TREC as two sets of activites
the main activity (core in TREC jargon)
subsidiary activities(tracks in TREC jargon)
the core has two types of tasks
Ad-hoc (that corresponds to retrospective retrieval)
Routing (that corresponds to the selective dissemination of information)
Relevance judgement for the old collections and old topic selection are made available in each subsequent TREC series experiment. The routing task involves using some of the old topics on the new collections: the relevance information from the old collection may be use to help formulate the query or the profile.
Out put list from each participating research team are sent to the NIST. Where they are merged for evaluation for each topic the hundred top ranking documents from all the participating teams are merged into a signal set which is then given to the assessor for relevance evaluation. This method called pooling.
First TREC was stated at 1992. From 1992-2005 it held 14th conferences in November month each and every year except TREC- 2. These conferences are co sponsored by DARPA&NIST.The 15th TREC conference will started in the month of November in this year
Benefits of TREC
Boolean retrieval
passage or paragraph retrieval
combining the results of more than one search
retrieval based on prior relevance assessment
query expansion & query reduction
String & concept based searching.
Dictionary based stemming. And so on.
Tuesday, March 21, 2006
Life
Life is beautiful.Beauty is life.You are life.For you are beautiful. |
Sunday, March 19, 2006
writing
| Is internet an information retrieval system? Internet is a kind of IRS [Information Retrieval System] it means the system should have information and it give the essential information for it's user{information need}. Internet has lot of information, websites are having that information's the library documents are having the information. An internet has lot of information and that would be indexed, because every IRS that should be indexed internet is also has be a kind of index but it will not appears on screen. An internet has some retrieve techniques. To day most of people aware of an internet because internet is a one of the popular. Users express their information need through the query an internet, it is key word search. Expressing way is important because retrying information depending upon query then system compares the query and indexed objects (database document) after the system retrieve the relevant information. Most of people fall in internet because it possible to retrieve the quick and latest information Internet-components 1. the documents (websites) 2. the index 3. the users (information need) 4. the query 5. search 6. comparison 7. database website Information retrieval system should have documents, index information need and its _expression and retrieve the relevant information than that would be called an IRS. So by all these aspects we can say internet is an information retrival system For example take a google ,if we want the information about" birds" then we will express our information need through the query because internet is a key word search .So we can type the word "birds" after internet search the birds and compare exist websites and internet display the number of available websites but it retrieve the starting 1000 websites only. So we can say an internet is an information retrieval system. |
Tuesday, March 07, 2006
Wednesday, February 01, 2006
ASSIGNMENT
Information retrieval process
comparison of 1&2 model
1.In 1st model we see the population of documents and then selected documents, but
in 2nd model we see only text objects means selected documents only.
2.1st model have a sysem vocabulary but it is not in 2nd model.
3.Users can expres is need through the request in 1st model. But in 2nd model user can
expres is information need through the query .
4.2nd model compared the users query and exist information. but it is not present in 1st
model .
5. when the retrieving information can't relevant to the users information need then it will be
feedback and altered the query or change the information need in 2nd model ,but in 1st
model not is.
6.users can be retrieval the information through the index or location of the information[when they are know about that location] in 1st model. but it is not in 2nd model.
comparison of 1&2 model
1.In 1st model we see the population of documents and then selected documents, but
in 2nd model we see only text objects means selected documents only.
2.1st model have a sysem vocabulary but it is not in 2nd model.
3.Users can expres is need through the request in 1st model. But in 2nd model user can
expres is information need through the query .
4.2nd model compared the users query and exist information. but it is not present in 1st
model .
5. when the retrieving information can't relevant to the users information need then it will be
feedback and altered the query or change the information need in 2nd model ,but in 1st
model not is.
6.users can be retrieval the information through the index or location of the information[when they are know about that location] in 1st model. but it is not in 2nd model.
Tuesday, November 29, 2005
Canon of relevance
| Canon of relevance Canon of relevance is enunciated thus: “A characteristics used as the basis for the classification of a universe should be relevant to the purpose of the classification”. This definition as given by Ranganathan. Sayers also given the definition for canon of relevance thus: “Each characteristics should be relevant to the purpose of the classification.” The canon of relevance says to used the essential characteristics, essential characterristics means: “Chosen characteristics which is that most useful for the purpose of the scheme, is called the essential characteristic.” EXAMPLE: - 1) Let us take the universe of the boys in a class-room. a) Let the purpose of classification be to divide the boys into convenient graded groups for tutorial work. Then mother tongue, intelligence, and extent of knowledge are relevant characteristics; but height, colour and physical strength are not. b) Let the purpose of classification be to divide the boys into convenient graded groups for physical game. Then height, physical strength and age are relevent characteristics; but colour , extent of knowledge and clothe are not. 2)Let us take the universe of books. a) Let the purpose of classification be to suit the needs of printers. Then typography, leading, margin, illustrations, and paper are relevant characteristics. But subject, author are not relevant characteristics. b) Let the purpose of classification be to suit the needs of the readers in a library. Then subject, langguage, author and year of publication are relevant charecteristics; but the cover of books, the paper used in the books are not. Reference:-
|
Saturday, November 12, 2005
Friday, November 11, 2005
Thursday, November 10, 2005
HAI
| To day is friday. I have to blogge. "smart work is better then to hard work" |
Tuesday, November 08, 2005
MAN
| Man made a mony, mony not made a man, Today mony is a very important things in society. |
SAD
Today i am very sad, because nearer to my village my friend was died by attack of an Elephant.
AIM Aim is a very importent part of life, but i don't have any aim. |
Sunday, November 06, 2005
holiday
“Today I wakeup very late. & I am very bored, & don’t know why? one of my relative that is my Anti came to our home at 12:30. I spent my time with my Anti son up to 2 hours, by playing cricket, caram. Evening 4’O clock my Anti went back to home.
A Free Bird
I was a free flying bird.
In this busy large world.
Hoping to see the stars in the noon.
Thinking to remove the scales from the moon
A Free Bird
I was a free flying bird.
In this busy large world.
Hoping to see the stars in the noon.
Thinking to remove the scales from the moon
Saturday, November 05, 2005
sun rising
sun rising
Thursday, November 03, 2005
To day i feel very bored by my journey incidence since morning. I waiting to catch the bus at 9'o' clock,i wont get any bus and at 9 i get bus in that bus condoctor told me that 'pass is not allowed' then i was shocked. Then i quarreled with him and make to allow the pass. winner v/s loser When a winner makes a mistake he says"i was wrong". When a loser makes a mistake, he says "I wasn't my fault". |
Friday, October 21, 2005
Subscribe to:
Posts (Atom)