12 ms·
Stract: Open-souce, non-profit search engine
- com 3y agoFast, feels clean and uncluttered to use and the search results are fairly high quality. I like the “optic” idea. After reading the about page, I’m not sure what the developers are trying to achieve? Perhaps a sort of alternative-Universe Google search funded by search-context AdWords?
- spaduf 3y agoReally like the explore feature. It lets you put in a url and shows you similar sites. Very promising project. Love to see people actually thinking about what search would be rather than rehashing decades old ideas.
- gregw134 3y agoWanted to say congrats on launching! I'm building a search engine myself, I can tell a lot of work went into this. I think the biggest thing you overlooked are page titles. When you issue a query it's a bit hard to quickly scan and judge what a site is about because the page titles are missing.
- pooper 3y agoHow do you crawl the web? Do you follow links around? How do you reach a page that isn't linked from anywhere you've crawled?
- gregw134 3y agoI'm just using common crawl for now
- mmkos 3y agoI mean that's what web crawling is, right? By extension, you just can't reach a page unless you stumble upon a link to it _somewhere_. Google gives you an option to submit a link and schedule a crawl that way, so that's another option if it's not being linked to from anywhere.
- vanous 3y agoCongrats! I tried to search for a particular domain data but neither search nor the explore would have the domain listed. What's the process to get unlisted domains indexed?
- daniel_iversen 3y agoawhh I can see DMOZ (https://en.wikipedia.org/wiki/DMOZ https://en.wikipedia.org/wiki/DMOZ) is no longer! That used to be the seed for crawling the internet I believe, for search engines.
- cristoperb 3y agoA static version (archive) of DMOZ is still available at http://www.odp.org/ http://www.odp.org/
- charcircuit 3y ago>how many bits are in a byte I checked 11 pages and none of the results were relevant.
- buo 3y agoI searched instead for "size byte bits", third result has the answer. It seems like the engine gives equal weight to all words in the search, so "are", "in" and "a" throw it off.
- kiwidrew 3y agoexcellent! I'm tired of search engines that optimize for natural language queries because the inevitable trade-off is that they become useless at keyword/exact queries.
- crotchfire 3y agoWhere does their crawl come from?
- kuratkull 3y agoIt's failing (completely wrong results) my goto query for testing search engines: "best sub 10 usd Linux single board computer" Try it out
- Levitating 3y agoDamn Pine64 has some fun stuff happening. Also I noticed DuckDuckGo performed much better than Google with this benchmark.
- jqpabc123 3y agoclearly labelled, contextual ads based on your current search query and a subscription option without ads Perfect! This is the way the god of the internet intended search engines to work. But DuckDuckGo does the same and currently provides superior results based on a very brief test. So good luck with that.
- godzillabrennus 3y agoDDG is amazing. To good to be viable I’m usually thinking.
- Brian_K_White 3y agoDDG sucked hard enough I'm now paying for kagi.
- notoverthere 3y agoDuckDuckGo also allows you to switch off ads, for free, without any fuss or adblocker needed. Just go to the settings page. Although if you aren't going to support DDG with ad revenue, I'd suggest supporting with a donation if you can afford it and value their service.
- jqpabc123 3y agoI really don't mind helping DDG take advertisers for all they are worth as long as it doesn't cost me my privacy or waste too much of my time. And if they take something away from Google in the process --- that's just an extra bonus. Turnabout is fair play don't you think? Google has worked very hard to take privacy away from users.
- candiddevmike 3y agoWhat would happen to DDG if Microsoft stopped letting them use Bing? What does MS get out of this relationship?
- jqpabc123 3y ago
- lastdong 3y agoOptics are a great idea, something we don’t see on other engines. Fully open source -`ღ´- Haven’t dig in to see what’s powering the search, I think DDG uses Bing
- Pufferbo 3y agoTried searching for Dota (the video game), and the game’s website is buried by a bunch of SEO spam. It might not even have been crawled because it doesn’t appear on the first or second page.
- AlienRobot 3y agoI searched for "horror movies" and the first result was a lemmy community that has literally "616 subscribers" "30 Posts" and "76 Comments" which is about as dead as you would expect from a lemmy instance. I also searched for "league of legends", and it couldn't find its homepage. I think its ranking algorithm may need improvement. Edit: also, I'd rather not say this, but do we really need another DuckDuckGo? I don't think Google fails at its job because of financial incentives. I think it might fail at this job simply put because the web of 2024 isn't the web of 1990. For example, the lemmy result, it's a link aggregation about horror movie articles. The search engine could literally do the job of the link aggregator, as it has a SERP that aggregates links, and yet it's aggregating links to link aggregators. Why are the search engines doing this? Because it's 2024. I wish someone tried a new approach at this problem rather than just copying Google's design and saying "it's Google but not yucky".
- redder23 3y ago[flagged]
- AlienRobot 3y agoThe reason I'd rather not say it is exactly because I don't want to sound like I'm just coming in here and shitting on things. I think there are no open source search engines because the instant you realize you have to periodically scrape the trillions of web pages in your index for updates you just give up because there is no way you can afford that without a solid business plan, which is hard to have when you are a search engine with no users, because you have no results, because you can't scrape trillions of webpages to build an index. Hence, I don't think it makes sense to try to make a general-purpose search engine, specially as Google has mastered that art and Google results look like that.
- skeptrune 3y agoThe option to return from only sites popular on HN, blogroll, and the other "manage optics" settings are incredibly cool and useful, I could see myself using this just for that feature alone. Exciting stuff.
- 3abiton 3y agoI'm more pessimistic about how would that drive bad actors to HN polluting the site.
- giancarlostoro 3y agoI take it you aren't showing dead threads? If you look at newer submissions you'll see people voting bad stuff to death. HN is insanely decent at community driven self moderation. Not to knock the mods who put a lot of work into the site of course, but I assume the community's own self-moderation helps some.
- sangnoir 3y ago"Polluting HN" is more than just the vitriol thats flagged - there are plenty of self-promoting, downright wrong, or comments clearly unrelated to the posted article but comment author gets to segue to their hobby-horse (off-topic discussions are annoyingly frequent, IMO). HN is better than most, but its not immune to being gamed. Once there is a financial incentive for it, it will become more common - see how Twitter turned out after offering monetary incentives for engagements.
- Fnoord 3y agoI mean, is your post off-topic? It is a tangent. Tangents are cool. You may or may not like one, that is OK.
- sangnoir 3y ago> I mean, is your post off-topic? My comment directly answered parent comment on an issue pertaining to search engines. How is that a tangent or remotely off-topic? This meta-discussion, on the other hand...
- PixelForg 3y agoSearch still needs some improvement, I typed "gundam watch order reddit" and was expecting some reddit links, but none of the results are reddit links. Perhaps there's another way to limit search results to a particular site here?
- mdhen 3y agoNormal way is "site:reddit.com"
- Gualdrapo 3y agoThat is Google way, rather than "normal"
- mdhen 3y agoI dont use google, that's how it works on ddg and kagi
- badsectoracula 3y agoThe site: operator seems to work in most search engines these days.
- logicprog 3y agoStract's githib page says they support site: queries
- jzelinskie 3y agoIf you're looking for the answer to this question the "relation graph" from AniDB is probably the best thing around: https://anidb.net/anime/715/relation/graph https://anidb.net/anime/715/relation/graph
- a1o 3y agoSearching for "adventure game studio", neither the website that has the forums or the GitHub repository is in the first page. Most results on the first page of search are really old things. Neither Wikipedia or repology that has the package infos are anywhere in the results.
- denysvitali 3y agoEveryone here is complaining about the search results - but instead I think we should all take some time to appreciate that someone worked hard to create a search engine (including the scraper / crawler part) and making it open source (AGPL). The results will be improved over time I guess, and for the few search queries I've done - I'm fairly happy with the results. Kudos to the authors!
- djbusby 3y agoI searched for a few things that were of the class "you should have this and match DDG/Kagi/G/Y/B" and the top 2/3 were matching. That's pretty good for a new-ish player.
- forgotmypw17 3y agoI've been doing some extensive searching for a particular topic for months using primarily Google, and I just found a bunch of sites that I had not previously found just running one query on this. I think that as with ChatGPT vs Bard, the result space is so huge, there are going to be many strength/weakness tradeoffs for any given query.
- Gratuity1901 3y agoIt is super hard to match the 'big' players. But straight programming a proper index listing is hard.
- ramon156 3y ago> type google > get anything but google ??
- viraptor 3y agoIt's an interesting case. The kneejerk reaction is "it should return google.com". But really... why? If I wanted google.com, I'd add the .com. If I wanted to search for something I wouldn't search for google first. I guess my top candidates would be: wikipedia page about google, google stock chart, recent news about google. google.com would never be a result I want to click. The current results are not amazing, but also do we care what the results for that one are? It's like someone telling you "cow" and expecting you'll know the context of what they're thinking of at the moment. Maybe a heading like "I have no idea what you're on about, here are some clickable ideas: google news, google stock price, ..." would be the best solution?
- lpellis 3y agoSeems surprising ok for coding related queries ('celery rate limit'), I'm curious about their scraping setup, building that out must be quite a big task.
- mcny 3y agoIn swagger/open API, why is everything a post? I tried the first endpoint get suggestions and tried searching for Gemini or Gemin hoping it would at least auto complete a word but the result set is empty. https://stract.com/beta/api/docs/#/autosuggest/route https://stract.com/beta/api/docs/#/autosuggest/route
- snvzz 3y agoThe search bar should really be full width. It can be very annoying to have your query not fit it while the window has plenty of room left.
- abrowniejr 3y agoI searched for "calories in 450 gm of steak" and the top 3 results were: 1. Brexit as the start of the reversal of neoliberal globalization - softpanorama.org 2. Directory Search - Fulshear-Katy Area Chamber of Commerce - chamberorganizer.com 3. The 100 Best New Products of 2020 - gearpatrol.com And none of the Page 1 results were related to my search query...
- keekslearns 3y ago[dead]
- stainablesteel 3y agothis is a neat thing, i like it, i'll add it to my list of search engines i use
- RDaneel0livaw 3y agoSo is this truly its own search engine / crawler / etc... and not using anyone else's searchs? I know ddg / kagi often use results from bing and other places, so just want to make sure. also, how can I add this to my firefox search inside the address bar / search field?
- fabrice_d 3y ago> also, how can I add this to my firefox search inside the address bar / search field? Navigate to https://stract.com/ https://stract.com/ then focus the url field: Firefox will display the new search engine at the bottom of the suggestions, on the "This time, search with:" line.
- mcdonje 3y agoI didn't see that option on mobile, but I got it added. Click the search provider icon in the search bar, go to search settings, then manage search providers, then add new. Add it with this url: https://stract.com/search?q=%s https://stract.com/search?q=%s
- dubbel 3y agoThat was hard to discover, thanks for your explanation. I was looking in "Firefox Settings -> Search -> Search Shortcuts" for a way to add it. I guess the functionality is not used very often, but it would be nice to have a hint on how to add new Search Engines there.
- RDaneel0livaw 3y agothank you, wow that's buried deep. Been using firefox for forever, and don't think I've ever noticed / seen that button. Thank you.
- tortoise_in 3y agoSo I have put two inquires of my local country but they didn't shown up
- bbsz 3y agoI think a lot of people will now go and benchmark queries only to report back disappointed with results. Trying to build generalized search engine for the modern internet that will come close to Google/Bing would require a "tech megaproject" level of investment and commitment. Most likely only to end up with the same optimizations and architecture as existing big-search and the very similar level of experience. I think it's a better direction to build a search based on more limited amount of topic-based data and focus on great match engine within, then - just aggregate the relevant ones together. Far more maintainable also on the crawling part. I can use google/bing to find the Honda dealership or read keyboard reviews, or get 50 most useful unix commands. I also wonder if with the rise of LLMs, while it still may not be feasible in such large scale production environment, those can serve as guides/agents to also improve the query itself and not the results of the query, for example - a chat-like search where user answers shift the relevancy metrics for returned documents. This would fit perfectly for smaller but open source, customizable and thematic search. That being said. I think it's great that project as such pop up more often. (Phind.com was also on my radar this year)
- safety1st 3y agoGenerates some pretty interesting results. No way to make it my default search engine?
- logicprog 3y agoOn Firefox just add its search URL (https://stract.com/search?q=%s https://stract.com/search?q=%s I think, it's somewhere above in this thread) to your list of search engines then set it as default. Idk about chrome
- safety1st 3y agoThank you. I spent 5 minutes and could not find a way to add this URL in Firefox. I'm just saying, I like the search engine, I hope that they would like to have users. I would like to be one of those users but have no clue how to add them to my browser so until that's fixed they're basically a non-starter.
- john-radio 3y agoVERY cool product. I have a quibble. I searched for "cool pokemon to use" and the top result was "How to use Paypal on Amazon" from "online-tech-tips.com". Understandable that the search results are not perfect - the second result was a perfect match for what I searched for - but anyway, clicking the "dials" icon gets me the following options: """Do you like results from online-tech-tips.com? (thumbs down, thumbs up, or banned emoji options) <a href='make-an-llm-do-something-stupid.com">Summarize result</a> """ IMO this feedback widget and (maybe) its backing API could use work. It's not that I like or don't like results from online-tech-tips.com; it's that they're a bad result for the specific context of this search.
- remram 3y agoSources: https://github.com/StractOrg/stract https://github.com/StractOrg/stract Backend in Rust (axum web framework, rocksdb), frontend with Svelte.
- vladstudio 3y agoTo make Stract usable for me (slightly reduced vision), I had to apply the following custom CSS: ``` html, body, div, td, th, p, h1, h2, h3, h4, h5, b, i, strong, li, button { font-family: ui-sans-serif, sans-serif !important; webkit-font-smoothing: antialiased; font-weight: 400; text-rendering: geometricPrecision; } ```
- eviks 3y ago> immensely customizable -- We aim to give you the ability to customize everything about the search. You can block sites, boost sites, prioritize links from specific sites and much, much more. Great! Can I use more than one optic? The drop-down list seems to allow only 1. > Oh, and if we ever become evil (maybe by changing our motto) please take our code and start a competitor. The most important part is the index data, what would be the deal with that?
- sydbarrett74 3y agoVery impressive, and kudos to the developers and originators. I just hope Stract doesn't go 'corporate' the way DDG did. :(
- highmastdon 3y agoGreat stuff! Just want to mention, when I search for “ExpressLRS use uart on older f4 fcs” it gives me about 15 results, but only the first two are unique. The other 13 are a literal copy of the first, both in content and in URL. Probably best to filter for uniqueness
- lock-the-spock 3y agoWonderful project, congratulations! I love the speed, clean design, many options, multilingual results, overall very impressive!! Some quibbles/points to consider: * I can't find anything on the people/organisation behind, and can onl guess from the Terms that the team is based in DK. * Search results are broad and interesting, maybe a bit more weighting for the joint occurrence of terms would be great. * Developing a site weight over time might be interesting, maybe even with user votes. Currently minor and major sites appear all together and e.g. a search for "Donald" gives me an interesting ranking order that gives neither the most famous Donald's nor the most reliable sites firet (not problematic per se - my fault for entering an unclear search term) * There are some interesting result patterns, with often official sites quite low. For instance search for "EU" with some term like subsidy (in any of the languages I speak) gives me random project websites but nothing from any of the official EU websites, or "Microsoft 365" (sorry...) gives me no MS website. * Very minor but hopefully a very easy fix: at least on Firefox mobile there is no direct way to add the search to my search engines, I had to add it manually. For other engines I can long press.on the search field and then get the option. Great work, keep it up! I will certainly start using this :-)
- bboygravity 3y ago"maybe even with user votes. " I'm so mind-blown that this does not exist yet. Free ranking feedback (live training of the algo!) + better search results for everyone. win-win
- viraptor 3y ago> Free ranking feedback Free spam SEO ranking in practice. A spam site has 1000x the incentive to upvote its result than you have to downvote it. YaCy did a distributed index with filtering lists and you effectively had to keep a list of who you trust / your own filter.
- bomewish 3y agoKinda solved with accounts, and/or subs, and monitoring/fraud detection? Or just turn it off for product related searches but keep it on for information related?
- icar 3y agoThe only current search engine that I can use in my native language, Catalan, is Google. I can't wait for a project like this one to get good at that.
- pcblues 3y agoIf you are interested in setting up your own non-profit org marketplace or know someone who does, I made an example one using free tools (https://donate.pcblues.com/ https://donate.pcblues.com/) that costs me only about $10 USD per month to host the example because it is just a hosted linux VM and not Saas or software subscription based. I configured the VM myself and then "just" installed the software and configured it. It hasn't been down for ages. I only just remembered to check it. It does everything from merch and service websites to escrowed payment transactions, user reputation, etc.
- pcblues 3y agoIt's just an offer to communicate, not a business.
- vaicorinthians 3y agoI would give stract.org a shot tho
- fulmicoton 3y agoThat looks quite promising! Thank you for crediting tantivy in the github README, that's well appreciated! Ping me if I can help with anything.
- latentdeepspace 3y agoCan someone provide a bit of background how the crawling part works?
- gardnr 3y agoI just set it as my default search engine for a day. It's not quite there for my use cases. Can we help improve the search results?
- bomewish 3y agoThought of grabbing like a big chunk of the way back machine and having THAT in the index? There’s always so much good stuff that gets nuked, and being able to search across it properly would be potentially very interesting.
- Brian_K_White 3y agoI just set up a YaCy jail on my truenas box at home. It's a distributed p2p system. Haven't actually used it yet since I'm currently paying for kagi and it's good, and I only just set it up yesterday. But this just struck me, I just said 2 things there and this post is yet another, between kagi, yacy, and now stract, not just 3 different names but 3 different types of solution to a problem, and all seemingly actually viable, that have popped up recently after decades of no one really feeling like they needed anything else. I think something is changing.
- blinding-streak 3y agoSad to say this for a promising idea, but the search results are objectively terrible. If it wants to succeed, it needs to nail the primary use case.
- logicprog 3y agoI really think the existence of this project is a Good Thing. Massive kudos to the people working on it. Previously I was always disappointed that our only options seemed to be open source meta search engines and closed source search engines, with most of the latter being corporate surveillance calitalist hellscapes, or anonymized portals to the same (Google and Bing, Duck Duck Go and Startpage), with only Kagi being an exception, although still closed source. It seemed like a really uniquely bad landscape, given that in most other areas of software there are at least some FLOSS alternatives to proprietary platforms that actually implement their core functionality, regardless of their relative quality. Stract finally changes that, which means several good things to me. First, you get actual, real accountability for them to stick to their stated privacy goals. Second, you get the ability for a wide variety of people to contribute to and influence the project, and/or learn from the project how to do this stuff themselves. And finally, most importantly, since it's a full reimplementation from indexing up, it's an opportunity to innovate on and experiment with the fundamentals, instead of just rearranging deck chairs on the titanic like e.g. Startpage. Thats really great :)
- logicprog 3y agoA few suggestions: - search results seem to be somewhat case sensitive, which is a massive problem for me when for instance searching programming terms - in general the matching algorithm seems way too strict, only matching against the exact thing you entered, which makes it very difficult to profitably search for things where you don't have exact specific terms in mind, like perhaps computer errors or genres of things. I think a lot of people's problem with how liberally Google interprets your search results is not that it interprets them liberally necessarily, but that it doesn't respect the other options it provides for trying to match things more strictly. As long as you provide something like Google's quote mechanism and actually respect it, I feel like it would be a lot better to match things more liberally by default. Maybe some amount of fuzzy searching, and matching by synonyms. Also you could probably just use a dictionary of synonyms to do that instead of whatever statistical model Google is probably using, in order to ensure more predictable results. - as someone else somewhere in this thread mentioned, it seems like stuff like a, an, and the, are all matched against and waited equally to other words. This, especially combined with the fact that it only matches words exactly makes the search results feel way too brittle and unforgiving
- gettodachoppa 3y agoFantastic! This is what I've been waiting for for 10 years, since Google removed the feature: a search engine which realizes 99% of the time, I want to search Discussions, and gives me the option to only show those. (reddit, forums, mastodon etc). This cuts down the SEO crap by 99.99%. That said, the results aren't great, hopefully it's something that improves as they index more pages. For example reddit doesn't seem to be indexed, why not? It's a goldmine of user content (even if the frontpage is 99% astroturfed US neolib propaganda).
- MythTrashcan 3y agoThis is a very cool search engine. Still suprised why this was made though. I thought there was a lot of other search engines already around that werw open-source. Anyways, interesting in seeing what changes as time goes on.
- Zuiii 3y agoSupports negative prompts! Will happily switch to this if I can figure out how to add it to firefox.
- aabbcc1241 3y agoIt has the same problem I experienced on Google search: When I search my projects it doesn't found it even when I use exact wording, adding github, npm, and username into search query doesn't help...