|
2025-08-21
|
|
|
Spot - Geospatial search for OpenStreetMap
Spot allows you to search for relations between entities in OpenStreetMap.
Spot can analyse your prompt for entities, their relations and location and display the matching results on a map.
DW Innovation
|
2024-08-18
|
|
|
TrainConnections
Search trains with TrainConnections. Plan and search international train routes effortlessly and discover the most convenient train routes across Europe or nearby regions. TrainConnections searches thousands of train routes to help you find and plan your trip quickly and easily. Wherever you want to travel, use our comprehensive train map to find eco-friendly travel options tailored to your needs.
|
2024-07-19
|
|
|
Maps Mania: AI Satellite Search
|
2021-08-10
|
|
|
Overpass API – OpenStreetMap Wiki
The Overpass API (formerly known as OSM Server Side Scripting, or OSM3S before 2011) is a read-only API that serves up custom selected parts of the OSM map data. It acts as a database over the web: the client sends a query to the API and gets back the data set that corresponds to the query.
Unlike the main API, which is optimized for editing, Overpass API is optimized for data consumers that need a few elements within a glimpse or up to roughly 10 million elements in some minutes, both selected by search criteria like e.g. location, type of objects, tag properties, proximity, or combinations of them. It acts as a database backend for various services.
|
2021-04-14
|
|
|
SEARCH RECORD
Tools for exploring your Google Search archive. A project by Jon Packles
|
2021-02-27
|
|
|
Flim
Glim is an AI-based search engine that allows searching for specific scenes in movies.
|
2021-02-14
|
|
|
Same Energy | Visual Search Engine
Find beautiful images.
|
2020-02-14
|
|
|
Google Dataset Search
|
2020-02-14
|
|
|
Travel And Mapping Software | Travel Time Map
Travel and mapping software to draw a travel time map. Draw a driving time map, export travel time catchment data or put the API on a site search feature.
|
2019-10-23
|
|
|
Open Knowledge Maps - A visual interface to the world's scientific knowledge
Start your literature search here: get an overview of a research topic, find relevant papers, and identify important concepts.
|
2019-08-06
|
|
|
GeoSeer - The spatial data search engine
GeoSeer is a search engine for spatial data. You can use it to find WMS, WFS, WCS, and WMTS datasets.
|
2019-07-23
|
|
|
Elasticsearch Library Search UI
Mit der Library Search UI können Entwickler Nutzerschnittstellen für die Suche nicht nur mit Elasticsearch erstellen.
|
2019-05-07
|
|
|
CC Search
CC Search is a tool that allows openly licensed and public domain works to be discovered and used by everyone. Creative Commons, the nonprofit behind CC Search, is the maker of the CC licenses, used over 1.4 billion times to help creators share knowledge and creativity online.
CC Search searches across more than 300 million images from open APIs and the Common Crawl dataset. It goes beyond simple search to aggregate results across multiple public repositories into a single catalog, and facilitates reuse through features like machine-generated tags and one-click attribution.
Creative Commons
|
2018-10-04
|
|
|
Making it easier to discover datasets
Similar to how Google Scholar works, Dataset Search lets you find datasets wherever they’re hosted, whether it’s a publisher's site, a digital library, or an author's personal web page.
|
2018-02-01
|
|
|
WhatIsWhere
Powerful map based POI search taking into account all your criteria to help you make the best decision for you or your business.
|
2017-04-27
|
|
|
Using SVG with CSS3 and HTML5, 1st Edition [Book]
Unlike other image formats, SVG can be an interactive part of your web site, integrated in HTML5 markup or created dynamically using JavaScript. It can be styled by CSS to instantly adapt to changes in your site's design. Flexible metadata options can make the graphics accessible to screen readers and search engines. And it is always displayed at the highest resolution possible, without requiring larger file sizes, making it ideal for both widescreen desktop displays and Retina mobile devices.
Using SVG with CSS3 and HTML5 starts with the basics, explaining how simple shapes and icons are defined in SVG, and then builds upon that foundation to create complex graphics and interactive, animated applications. It covers all the features you're likely to come across in web design, while avoiding areas with poor browser support.
|
2017-04-26
|
|
|
Google takes steps to limit offensive and inaccurate search results - The Verge
Google announced a new set of changes to its search engine, in an effort to deliver higher quality results and limit fake news. The company outlined changes to its search ranking, feedback...
|
2016-03-03
|
|
|
How We Monitor and Run Elasticsearch at Scale - SignalFx
|
2015-08-31
|
|
|
Handpicked free fonts for graphic designers with commercial-use licenses
Font Squirrel scours the internet in search of FREE, highest-quality, designer-friendly, commercial-use fonts and presents them for easy downloading. We don't have the most, but we do have the best.
|
2015-05-18
|
|
|
Export Google Search History
|
2015-03-03
|
|
|
Global Trend Map
Interactive 3D visualisiation of global Google search trends
|
2015-01-12
|
|
|
Google Search Appliance - Datenblatt
|
2014-05-30
|
|
|
Google Maps Mania: How to Find Old Maps Online
|
2012-07-15
|
|
|
Pattern | CLiPS
Pattern is a web mining module for the Python programming language.
It bundles tools for data retrieval (Google + Twitter + Wikipedia API, web spider, HTML DOM parser), text analysis (rule-based shallow parser, WordNet interface, syntactical + semantical n-gram search algorithm, tf-idf + cosine similarity + LSA metrics), clustering and classification (k-means, KNN, SVM), and data visualization (graph networks).
|
2012-07-11
|
|
|
Large-scale Incremental Processing Using Distributed Transactions and Notifications
Updating an index of the web as documents are crawled requires continuously transforming a large repository of existing documents as new documents arrive. This task is one example of a class of data processing tasks that transform a large repository of data via small, independent mutations. These tasks lie in a gap between the capabilities of existing infrastructure. Databases do not meet the storage or throughput requirements of these tasks: Google's indexing system stores tens of petabytes of data and processes billions of updates per day on thousands of machines. MapReduce and other batch-processing systems cannot process small updates individually as they rely on creating large batches for efficiency.
We have built Percolator, a system for incrementally processing updates to a large data set, and deployed it to create the Google web search index. By replacing a batch-based indexing system with an indexing system based on incremental processing using Percolator, we process the same number of documents per day, while reducing the average age of documents in Google search results by 50%.
Daniel Peng, Frank Dabek
|
2012-04-24
|
|
|
ConceptNet 5
ConceptNet is a semantic network containing lots of things computers should know about the world, especially when understanding text written by people.
It is built from nodes representing concepts, in the form of words or short phrases of natural language, and labeled relationships between them. These are the kinds of things computers need to know to search for information better, answer questions, and understand people's goals. If you wanted to build your own Watson, this should be a good place to start!
Attribution-ShareAlike License (CC BY-SA)
|
2012-01-11
|
|
|
LOCATING LONDON'S PAST
This website allows you to search a wide body of digital resources relating to early modern and eighteenth-century London, and to map the results on to a fully GIS compliant version of John Rocque's 1746 map.
Locating London's Past
|
2011-12-01
|
|
|
An Enhanced Indexing And Ranking Technique On The Semantic Web
With the fast growth of the Internet, more and more information is available on the Web. The Semantic Web has many features which cannot be handled by using the traditional search engines. It extracts metadata for each discovered Web documents in RDF or OWL formats, and computes relations between documents. We proposed a hybrid indexing and ranking technique for the Semantic Web which finds relevant documents and computes the similarity among a set of documents. First, it returns with the most related document from the repository of Semantic Web Documents (SWDs) by using a modified version of the ObjectRank technique. Then, it creates a sub-graph for the most related SWDs. Finally, It returns the hubs and authorities of these document by using the HITS algorithm. Our technique increases the quality of the results and decreases the execution time of processing the user's query.
Ahmed Tolba, Nabila Eladawi, Mohammed Elmogy
|
2011-12-01
|
|
|
Structure Learning of Probabilistic Graphical Models: A Comprehensive Survey
Probabilistic graphical models combine the graph theory and probability theory to give a multivariate statistical modeling. They provide a unified description of uncertainty using probability and complexity using the graphical model. Especially, graphical models provide the following several useful properties:
- Graphical models provide a simple and intuitive interpretation of the structures of probabilistic models. On the other hand, they can be used to design and motivate new models.
- Graphical models provide additional insights into the properties of the model, including the conditional independence properties.
- Complex computations which are required to perform inference and learning in sophisticated models can be expressed in terms of graphical manipulations, in which the underlying mathematical expressions are carried along implicitly.
The graphical models have been applied to a large number of fields, including bioinformatics, social science, control theory, image processing, marketing analysis, among others. However, structure learning for graphical models remains an open challenge, since one must cope with a combinatorial search over the space of all possible structures.
In this paper, we present a comprehensive survey of the existing structure learning algorithms.
Yang Zhou
|
2011-11-28
|
|
|
YaCy - Freie Suchmaschinensoftware und dezentrale Websuche
Dezentrale Web-Suche mit YaCy
|
2011-11-09
|
|
|
django-rdflib and postgresql - the best of both worlds — EuroPython 2011: Florence, June 20–26
rdflib is a python library implementing a database with various triples back-end, parser, data serializers, SPARQL is a Python interface to extract/insert triples. We integrated it in Django reusing the database connection and exposing an ORM interface, along with full-text search on literals. This presentation shows a django-rdflib case study with a PostgreSQL backend in Brain Architecture Management System (http://brancusi1.usc.edu) - a neuroscientific project for the University of Southern California. Benefits of the flexible RDF structure will be shown, allowing researchers to insert free format data, making data public with a customizable serialization and use the powerful full-text search integrated in PostgreSQL.
|
2011-11-07
|
|
|
CommonCrawl
Common Crawl produces and maintains a repository of web crawl data that is openly accessible to everyone. The crawl currently covers 5 billion pages and the repository includes valuable metadata. The crawl data is stored by Amazon’s S3 service, allowing it to be bulk downloaded as well as directly accessed for map-reduce processing in EC2. This makes wholesale extraction, transformation, and analysis of web data cheap and easy. Small startups or even individuals can now access high quality crawl data that was previously only available to large search engine corporations.
|
2011-10-26
|
|
|
Search VideoLectures for iswc - videolectures.net
Great page to get an overview on what's happening in the Semantic Web
|
2011-08-01
|
|
|
15 Tools to Help Speed Up Your Website
The speed at which a Website loads is paramount to maintaining a positive user experience and, as we learned last year, has a direct impact on the site's organic search rankings on Google. The search giant's recent beta launch of its Page Speed Service gives us the latest in a long line of products and tools designed to help site owners boost page load speed. In what is by no means a comprehensive list, we've outlined a few such tools worth checking out....
ReadWriteWeb
|
2011-07-08
|
|
|
bigdata | Download bigdata software for free at SourceForge.net
bigdata(R) is a scale-out storage and computing fabric supporting optional transactions, very high concurrency, and very high aggregate IO rates.
Features statement-level provenance, free-text search, and incremental load and retraction, inference etc.
|
2011-07-04
|
|
|
Iconfinder
Search through 155,796 icons or browse 812 icon sets
|
2011-06-10
|
|
|
Google, Microsoft, and Yahoo Team Up to Advance Semantic Web - Technology Review
A push to add meaning to Web pages to aid search could also enable other kinds of intelligent web apps.
Tom Simonite
|
2011-06-10
|
|
|
The False Choice of Schema.org | The Beautiful, Tormented Machine
Some of you may have heard that Microsoft, Google and Yahoo have just released a new uber-vocabulary for the Web. As the site explains, if you use schema.org, you will get a better looking search listing on all of the search listings for Bing, Google and Yahoo. While this may sound good on the surface, it is very bad news for choice on the Web. There are few points that I’d like to make in this post:
Manu Sporny
|
2011-06-06
|
|
|
BeyondTrees and the New York Times: Using Lucene to build a time machine
Imagine you can see 160 years of history, all on one screen. You can zoom and pan, you can look at a particular day, you can even do a search. And when you do, the results come up not as a list, but as a heat map that shows where in history that topic appears, and how often.
|
2011-05-11
|
|
|
RDF JSON - docs.api
User documentation for the Talis Platform APIs.
The Talis Platform is a Web based environment for building Semantic Web applications and services. It is a hosted system which provides an efficient, robust storage infrastructure for documents and metadata. The Platform is accessed using a suite of Web based services which provide sophisticated data management, query, indexing and search features.
http://www.talis.com/platform
|
2011-04-19
|
|
|
kd-tree - Wikipedia, the free encyclopedia
In computer science, a kd-tree (short for k-dimensional tree) is a space-partitioning data structure for organizing points in a k-dimensional space. kd-trees are a useful data structure for several applications, such as searches involving a multidimensional search key (e.g. range searches and nearest neighbour searches).
|
2011-04-15
|
|
|
Accepted Papers -- W3C Workshop on Web Tracking and User Privacy
This workshop serves to establish a common view on possible Recommendation-track work in the Web privacy and tracking protection space at W3C, and on the coordination needs for such work.
The workshop is expected to attract a broad set of stakeholders, including implementers from the mobile and desktop space, large and small content delivery providers, advertisement networks, search engines, policy and privacy experts, experts in consumer protection, and other parties with an interest in Web tracking technologies, including the developers and operators of Services on the Web that make use of tracking technologies for purposes other than to behavioral advertising.
|
2011-01-12
|
|
|
GeoNetwork opensource
"GeoNetwork is a catalog application to manage spatially referenced resources. It provides powerful metadata editing and search functions as well as an embedded interactive web map viewer. It is currently used in numerous Spatial Data Infrastructure initiatives across the world."
Implements INSPIRE - integrates AGROVOC?
|
2010-11-23
|
|
|
RDF Data Analysis with Activation Patterns
RDF data can be analyzed with various query languages such as SPARQL
or SeRQL. Due to their nature these query languages do not support fuzzy queries.
In this paper we present a new method that transforms the information presented
by subject-relation-object relations within RDF data into Activation Patterns. These
patterns represent a common model that is the basis for a number of sophisticated
analysis methods such as semantic relation analysis, semantic search queries, unsuper-
vised clustering, supervised learning or anomaly detection. In this paper, we explain
the Activation Patterns concept and apply it to an RDF representation of the well
known CIA World Factbook.
Peter Teufl and Günther Lackner
|
2010-11-12
|
|
|
A SURVEY OF EIGENVECTOR METHODS FOR WEB INFORMATION RETRIEVAL
Web information retrieval is significantly more challenging than traditional well controlled,
small document collection information retrieval. One main difference between traditional information retrieval and Web information retrieval is the Web's hyperlink structure. This structure has been exploited by several of today's leading Web search engines, particularly Google and Teoma. In this survey paper, we focus on Web information retrieval methods that use eigenvector
computations, presenting the three popular methods of HITS, PageRank, and SALSA.
AMY N. LANGVILLE, CARL D. MEYER
|
2010-11-09
|
|
|
A Node Indexing Scheme for Web Entity Retrieval
Now motivated also by the partial support of major search engines, hundreds of millions of documents are being published on the web embedding semi-structured data in RDF, RDFa and Microformats. This scenario calls for novel information search systems which provide effective means of retrieving relevant semi-structured information. In this paper, we present an entity retrieval system designed to provide entity search capabilities over datasets as large as the entire Web of Data. Our system supports full-text search, semi-structural queries and top-k query results while exhibiting a concise index and efficient incremental updates. We advocate the use of a node indexing scheme and show that it offers a good compromise between query expressiveness, query processing time and update complexity in comparison to three other indexing techniques. We then demonstrate how such system can effectively answer queries over 10 billion triples on a single commodity machine.
Renaud Delbru, Nickolai Toupikov, Michele Catasta, Giovanni Tummarello
|
2010-11-09
|
|
|
Using Naming Authority to Rank Data and Ontologies for Web Search
The focus of web search is moving away from returning relevant documents towards returning structured data as results to user queries. A vital part in the architecture of search engines are link-based ranking algorithms, which however are targeted towards hypertext documents. Existing ranking algorithms for structured data, on the other hand, require manual input of a domain expert and are thus not applicable in
cases where data integrated from a large number of sources exhibits enormous variance in vocabularies used. In such environments, the authority of data sources is an important signal that the ranking algorithm has to take into account. This paper presents algorithms for prioritising data returned by queries over web datasets expressed in RDF. We introduce the notion of naming authority which provides a correspondence between identifiers and the sources which can speak authoritatively for these identifiers. Our algorithm uses the original PageRank method to assign authority values to data sources based on a naming authority graph, and then propagates the authority values to identifiers referenced in the sources. We conduct performance and quality evaluations of the method on a large web dataset. Our method is schema-independent, requires no manual input, and has applications in search, query processing, reasoning, and user interfaces over integrated datasets.
Andreas Harth, Sheila Kinsella, Stefan Decker
|
2010-09-27
|
|
|
Implementing LinkedIn Faceted Search
Nice description and linked to open-source projects from linkedin team.
|
2010-09-21
|
|
|
Faceted Wikipedia Search
|
2010-08-18
|
|
|
Friend of a friend (FOAF) search engine
A FOAF search engine.