“The Lumen database collects and analyzes legal complaints and requests for removal of online materials, helping Internet users to know their rights and understand the law. These data enable us to study the prevalence of legal threats and let Internet users see the source of content removals.” #censorship #search-engine
on 02026-08-12"Offshore Leaks" is a #search-engine for the Panama Papers, etc.
on 02026-08-12Boardreader is a #search-engine for forum threads, but it doesn’t seem to have very much information.
on 02026-08-12Vladimir Prelovac of "Kagi" thinks the age of #PageRank is over, so people need to pay them to use their #search-engine
on 02023-01-15the marginalia.nu #search-engine may be abandoned.
on 02022-05-21#search-engine (veneer over Google) to show the oldest surviving result on the web
on 02022-05-21spambots pounding on the searchmysite.net #search-engine for some reason? marginalia.nu reports that they don’t seem to pay attention to the results
on 02022-05-21“Prefix Sums and Their Applications”, Blelloch 93. This is the definitive #paper on #prefix-sum #algorithms as of uh 22 years ago. It mentions specifically a #regular-expression #search-engine and #lexing as two of the applications of prefix sum! I think the “parallel solution of recurrence problems” mentioned here can be applied to #DSP IIR filtering. This is just “Chapter 1” (of what, I have no idea), but it promises more meat in chapters 2, 3, and 4, mostly to do with linked lists and trees.
on 02015-08-15“#Levenshtein automata can be simple and fast” and useful for finding potential misspellings (like for a #search-engine, with applications given to #Lucene) in a #trie. #Python with #graphviz: 40 lines of code and good (O(max edit distance) supposedly) worst-case #complexity. #smallisbeautiful #algorithms
on 02015-08-15