ArticleslgStudy

computer science

List of SIMILE projects

List of SIMILE projects is a computer science topic covered in the lgStudy science library. This page brings together a partial reference excerpt, illustrations, worked examples, real-world applications and a short study plan, so you can understand List of SIMILE projects rather than just read about it. In short: The following is a list of SIMILE projects. The SIMILE tools assist in the storage, querying, transformation and mapping of very large collections of RDF data.

Key takeaways

  • List of SIMILE projects belongs to computer science; place it in that map before memorising details.
  • Learn the definition first, then one example that makes the definition concrete.
  • Connect List of SIMILE projects to a quantity you can measure, compute or draw — that is where exam questions come from.
  • Reproduce the core statement of List of SIMILE projects from memory before moving on to harder problems.

Reference excerpt

The following is a list of SIMILE projects. The SIMILE tools assist in the storage, querying, transformation and mapping of very large collections of RDF data. The tools developed within SIMILE are meant to allow people who are not Semantic Web developers to create ontologies which describe their specialized metadata, create RDF and convert other types of metadata into RDF. These open source tools are designed to be scalable and provide for cross-community sharing of metadata at low cost.

Longwell Longwell is a faceted browser which enables the user to visualize and browse any RDF data set, allowing the user to quickly build a user-friendly web site out of the RDF data without requiring the user to write any RDF code. Facets are metadata fields considered important for a given data set. In its default configuration, the collection of facets is returned along the right-hand side of the page, and clicking on any facet causes the refinement of facets in relation to the data retrieved. Longwell then displays only the subset of the data which meet those restrictions. This appears on the left-hand side of the page. Previously selected restrictions can be removed, which causes a broadening of the subset of items displayed.

Piggy Bank Piggy Bank is a Firefox extension which enables the user to collect information from the Web, save it for future use, tag it with keywords, search and browse information collected, retrieve saved information, share collected information and install screen scrapers. Piggy Bank gathers RDF data where it is available, and where it is not available, it generates it from HTML by using screen scrapers. This incremental approach to the realization of the Semantic Web vision allows the user to save and tag information gathered from web pages without having to cut, paste and label the various products of their browsing. By clicking on the keyword they have used to tag particular types of item, the user can view all of those items together within her browser, without having to open other applications. Users can also deposit saved data in the Semantic Bank, where other users can browse it and add their own contributions. This pooling of keywords underlies services such as Flickr and del.icio.us, where communities can collaborate to build a taxonomy for shared data. These taxonomies, which emerge as information is accumulated, are known as folksonomies.

Solvent Solvent is a Firefox extension that enables the user to write screen scrapers for Piggy Bank.

Gadget Gadget is an XML inspector which enables the user to condense large amounts of well-formed XML data.

Welkin Welkin is a graph-based RDF visualizer. It graphs RDF data sets, allowing the user to visualize the global shape and clustering characteristics of the data, which can aid them in mentally modeling it, seeing how it connects and identifying mappings between the set and possible ontologies. A particular data cluster which stands out when graphed might well be missed when browsed at closer range.

Fresnel Fresnel is a vocabulary for specifying how RDF graphs are presented. Fresnel addresses the problem that currently, each RDF browser and visualization tool decides, on an ad hoc basis, what information in an RDF graph is presented and how to present it. Fresnel uses the concepts of lenses and formats. Lenses determine which properties are displayed and how they are ordered. Formats control how resources and properties are presented.

Timeline Timeline is a tool for visualizing events over time. It can be populated by pointing it at an XML file

Exhibit Exhibit is technology that enables developers to provide browsing of faceted classifications in a web browser.

Referee Referee is a program that crawls the links that point to its user's pages. It extracts metadata from those pages and the text around the links that pointed to its user's pages, converting it, if need be, into RDF format. Referee discriminates between the pages that refer to the user's pages and the comments, meaning the text immediately surrounding the link. It generates a data graph, allowing it to display the fact that, for example, exactly the same comment in relation to its user's pages appears on more than one page, which is the container of the comment. A page can have more than one comment, and a comment can appear on more than one page. This can be illustrated in a data graph, but would not be possible with a data tree, such as is generated by the XML data model.

RDFizer The RDFizer project is a directory of tools for converting various data formats into RDF. MIT Libraries provides a home for some of these tools. RDFizers are a group of tools that allows the transformation of existing data into an RDF representation. Given a database of interest, these tools can often - when the data formats are highly structured -convert the data into an RDF representation without human intervention, first determining what ontology to use to express the information. Where semantic relationships are implicit, the RDFizers will not be as successful without human input. The SIMILE project has built RDFizers that convert from the following formats:

JPEG Joint Photographic Experts Group (Digital Photo-METADATA). MARC United States Library of Congress MAchine-Readable Cataloging of bibliographic data. MODS Metadata Object Description Schema for bibliographic element sets. OAI-PMH Open Archives Initiative Protocol for Metadata Harvesting. OCW Open Course Ware EMail BibTeX a tool for formatting lists of references usually associated with LaTeX documents. Flat Weather Java is an object-oriented applications programming language Javadoc tool for generating API documentation into HTML format from Java source code. Subversion or SVN is a software revision control system. Random

Crowbar Crowbar is a web scraping environment based on the use of a server-side headless Mozilla-based browser. It is used as a research prototype to investigate how to enable the running of Piggy Bank JavaScript scrapers from the command line and thus automate web site scraping.

References

Worked examples

Example 1 — a first encounter with List of SIMILE projects

Start with the simplest possible case. Write down what List of SIMILE projects claims or describes in one sentence, then invent the smallest concrete situation in which that sentence is true. In computer science, the smallest case is usually a single object, a single equation or a single measurement. Check that every symbol or term in your sentence has a meaning in that case.

Example 2 — changing one variable

Take the situation from Example 1 and change exactly one quantity: double it, halve it, or set it to zero. Predict what should happen to List of SIMILE projects before you calculate. Comparing your prediction with the result is the fastest way to find out whether you understand the idea or only the words.

Example 3 — an exam-style question

Typical questions about List of SIMILE projects ask you to (a) state it precisely, (b) apply it to given data, and (c) explain a limitation. Practise writing all three answers in under five minutes; the third part is what separates a full-mark answer from an average one.

Applications of List of SIMILE projects

In research
List of SIMILE projects appears in computer science research whenever the underlying quantities have to be modelled precisely. Papers usually cite it as a starting assumption and then explore where it breaks down.
In technology and industry
Engineering practice reuses List of SIMILE projects in design rules, simulations and safety margins. Knowing the idea lets you read a specification sheet and understand why the numbers look the way they do.
In the classroom
List of SIMILE projects is common in secondary-school and first-year university syllabi. It links to neighbouring topics Computing-related lists, Ontology (information science), Semantic Web, so understanding it makes those chapters shorter.
In everyday life
Look for List of SIMILE projects outside the textbook — in sport, cooking, traffic, electronics or the sky above you. An example you found yourself is remembered far longer than one you were given.

Affiliate

Preply — study more efficiently by working with a personal tutor. 50% off.

How to study List of SIMILE projects in 20 minutes

  1. Read the reference excerpt below once, without taking notes.
  2. Close the page and write down what List of SIMILE projects means in your own words.
  3. Compare your version with the excerpt and mark what you missed.
  4. Work through the three examples above with pen and paper.
  5. Explain List of SIMILE projects out loud to somebody else — or to Teacher Smith in the lgStudy chat.

Frequently asked questions

What is List of SIMILE projects in simple terms?

The following is a list of SIMILE projects. The SIMILE tools assist in the storage, querying, transformation and mapping of very large collections of RDF data.

Why does List of SIMILE projects matter?

Because it connects several computer science ideas at once: it gives you a definition you can apply, a quantity you can calculate, and a way to check whether a result is plausible.

How should I study List of SIMILE projects?

Read the excerpt, restate it from memory, then work through the examples and applications listed on this page. The five-step study plan above takes about twenty minutes.

What does this page cover?

It gives you a compact reference excerpt plus original lgStudy explanations, examples, applications and study material on List of SIMILE projects.

Tags

  • Computing-related lists
  • Ontology (information science)
  • Semantic Web

Keep exploring