How can I make my own version of IBM Watson?

Answer Wiki

3 Answers

Phillip Rhodes

Phillip Rhodes, worked at IBM

Answered Feb 2, 2014

Building a literal duplicate of Watson would be damn tough and cost a lot of money.  But you can do some similar things and get a start a few ways...

For an overview of how to build a "watson-like" you can start here:

https://www.ibm.com/developerwor...

For more on some of the Open Source software you would use to build something like Watson, look into NLP software, semantic knowledgebases, and reasoning software.  Some good starting points include:

http://opennlp.apache.org

http://uima.apache.org

GATE.ac.uk - index.html

Quepy: A Python framework to transform natural language questions to queries.

http://jena.apache.org

Watson also apparently used Prolog heavily, so you might want to look into Prolog resources. I cataloged a few here a  while back:

Prolog? I'm Going To Learn Prolog??

For some interesting research on interacting with natural language queries, see:

http://start.csail.mit.edu

And you might also look at another project that bills itself as "an open source wolfram alpha", here:

SymPy Gamma


Be prepared to invest quite a lot of time and energy.

4.1k Views · 21 Upvotes · View Timeline

Upvote21Downvote

Comments2+

Share

Promoted by VisionMobile

Are you coding for AR/VR?

Take this developer survey and let us know what is your primary coding language.

Start now at s.developereconomics.com

경축! 아무것도 안하여 에스천사게임즈가 새로운 모습으로 재오픈 하였습니다.
어린이용이며, 설치가 필요없는 브라우저 게임입니다.
https://s1004games.com

Gary Miller

Gary Miller, Software Developer, Business Intelligence Architect, AI Researcher

Answered Feb 10, 2015

This book contains a lot of the prerquisite knowledge needed to get started.

It covers the front end search portion.

Taming Text: How to Find, Organize, and Manipulate It: Grant S. Ingersoll, Thomas S. Morton, Andrew L. Farris: 9781933988382: Amazon.com: Books

I would follow this up though with resources on interfacing PROLOG with database using a Fuzzy model.

I haven't read this one yet due to its cost but going through the table of contents this one should be high on your list.

Knowledge Management in Fuzzy Databases (Studies in Fuzziness and Soft Computing): Olga Pons, Maria A. Vila: 9783540668367: Amazon.com: Books

You won't be able to get enough performance out of the high end PC to do a large knowledge base.  So you'll  probably want to plan on a high end PC cluster so that you can scale your solution.  Many existing programming languages do not do distributed comping without a lot of complex software infrastructure.  Since NOSQL  already provides you a way of scaling your knowledge out access multiple nodes as documents, I would recommend one of these as a scalable data storage mechanism.

1k Views · View Timeline

UpvoteDownvote

Comment

Share

Jack Park

Jack Park, Startup entrepreneur: knowledge gardens

Updated Feb 11, 2015

I think that's a question which is being answered in a variety of places. Actually, not an open source version of Watson in any literal sense, but there are open source projects, which include:

OAQA which is a student project actually associated with Watson
Watsonsim another student project
YodaQA -- I built and played with this one
AKSW OpenQA
AKSW QA
AquaLog
and my own OpenSherlock (was SolrSherlock).

Most of those are at GitHub or elsewhere on the web. OpenSherlock has not really found its way to GitHub, though early components of it are at GitHub under the old name SolrSherlock. Update: OpenSherlock now has its own repo: OpenSherlock

There are many papers and slides online about those various projects. One slidedeck for SolrSherlock is here
SolrSherlock: Linkfinding among Biomolecules with Literature-based Di…
and at the end of this month, a slidedeck for a conference talk about OpenSherlock will be at my slideshare account (root of the above link).

My experience with YodaQA is interesting: it's UIMA based, uses Solr for indexing Wikipedia, and uses Fuseki for indexing DbPedia and Yago, then answers questions.  It took me a while to bring it up on a Windows platform (used Cygwin for parts of it). In terms of question answering, it, and perhaps most of the open source so-called "cognitive systems" are not particularly accurate. I believe that's simply because a year or so of open source experiments is no match for the zillions of person-years invested by IBM in Watson. That's not to say open source won't rise to the occasion; I believe it will.

I did not mention:
OpenCog
DeepDive

earlier because they are not strictly question-answering platforms; but they each include powerful tools very useful in such a project; DeepDive does have an app written on it that approaches question answering.

1.8k Views · 5 Upvotes · View Timeline

 

[출처] https://www.quora.com/How-can-I-make-my-own-version-of-IBM-Watson

본 웹사이트는 광고를 포함하고 있습니다.
광고 클릭에서 발생하는 수익금은 모두 웹사이트 서버의 유지 및 관리, 그리고 기술 콘텐츠 향상을 위해 쓰여집니다.
번호 제목 글쓴이 날짜 조회 수
578 공공데이터 활용하기 - 002 file 가을의곰 2017.06.18 81
577 Learn the Watson API 가을의곰 2017.06.18 104
576 소프트웨어 | 카비레이크 윈도우7 설치하기~ file 가을의곰 2017.06.18 136
575 < 테샛 핵심 문항 70선 > file 졸리운_곰 2017.06.18 295
574 matlab 을 사용한 머신러닝 e북 file 졸리운_곰 2017.06.10 86
573 Squid 프록시 서버 가을의곰 2017.06.10 228
572 Varnish 이야기 [web proxy] 서버 file 가을의곰 2017.06.10 92
571 Expert System development with open source 졸리운_곰 2017.06.03 82
570 The PHLIPS Project : clips and php integration file 졸리운_곰 2017.05.31 65
569 Python Knowledge Engine (PyKE) file 졸리운_곰 2017.05.31 106
568 OpenNLP 퀵 가이드 file 졸리운_곰 2017.05.30 1184
» How can I make my own version of IBM Watson? file 졸리운_곰 2017.05.30 85
566 엑셀로 배우는 인공지능 file 졸리운_곰 2017.05.27 228
565 Why a microservices approach to building applications? file 졸리운_곰 2017.05.20 136
564 시벨리우스 무료강좌 사이트 졸리운_곰 2017.05.20 168
563 마이크로서비스로 리팩토링하기 3부: 마이크로서비스 도입을 위한 로드맵 file 졸리운_곰 2017.05.20 78
562 마이크로서비스로 리팩토링하기 2부: 데이터를 옮길 때 고려할 사항 file 졸리운_곰 2017.05.20 83
561 마이크로서비스로 리팩토링하기 1부: 마이크로서비스로 마이그레이션할 때의 고려할 것들 file 졸리운_곰 2017.05.20 242
560 기가막힌 "마이크로서비스" 관련 자료 링크 : Awesome Microservices file 졸리운_곰 2017.05.20 1426
559 browser based games written with PHP, MySQL, JavaScript, CSS, HTML etc file 졸리운_곰 2017.05.20 79
대표 김성준 주소 : 경기 용인 분당수지 U타워 등록번호 : 142-07-27414
통신판매업 신고 : 제2012-용인수지-0185호 출판업 신고 : 수지구청 제 123호 개인정보보호최고책임자 : 김성준 sjkim70@stechstar.com
대표전화 : 010-4589-2193 [fax] 02-6280-1294 COPYRIGHT(C) stechstar.com ALL RIGHTS RESERVED