![]()
A) Automated Integration B) Advanced Interface C) Analysis & Investigation D) Artificial Intelligence
A) Toyota B) Sony C) LG D) Honda
A) R2-D2 B) EVE C) BOLT D) C-3PO
A) Robot B-9 B) C-3PO C) WALL-E D) Optimus Prime
A) Snow Crash B) Neuromancer C) I, Robot D) Do Androids Dream of Electric Sheep?
A) Mother B) HAL 9000 C) Skynet D) Ultron
A) Bender B) Johnny 5 C) Dalek D) R2-D2
A) Anthropomorphism B) Algorithm C) Automation D) Articulation
A) Blue Origin B) Rethink Robotics C) Boston Dynamics D) iRobot
A) Deep Reinforcement Learning B) Genetic Algorithms C) Programming by Demonstration D) Artificial Neural Networks
A) Virtual Reality B) Social Networking C) Wireless Connectivity D) Machine Learning
A) Japan B) Germany C) China D) South Korea
A) Robots.txt B) Sitemap.xml C) CrawlerRules.json D) MetaTags.html
A) Vint Cerf B) Tim Berners-Lee C) Martijn Koster D) Charles Stross
A) CrawlerExclusion.txt B) RobotsNotWanted.txt C) PageAccessControl.txt D) WebBotRules.txt
A) Implementing CAPTCHA systems B) Using encryption C) Countering with security through obscurity D) Deploying firewalls
A) Courts mandate the creation of robots.txt files for all websites. B) Legal cases have shown that robots.txt is irrelevant to bot operations. C) Robots.txt is always ignored by courts in such cases. D) It has been used as a basis for legal action against non-compliant bot operators.
A) 500 B) 100 C) 50 D) 306
A) Institute of Electrical and Electronics Engineers (IEEE) B) World Wide Web Consortium (W3C) C) International Organization for Standardization (ISO) D) Internet Engineering Task Force (IETF)
A) Disallow B) Sitemap C) Crawl-delay D) Allow
A) Yes, if they are linked from another page that is crawled B) They only appear if the robots.txt file is missing C) Only if the site owner approves it manually D) No, they will never appear in search results
A) Crawl-delay B) Content-Signal C) Sitemap D) Disallow
A) All crawlers B) Googlebot C) Yandex D) BingBot
A) Facebook, Twitter, Instagram B) Amazon, eBay, Alibaba C) Ask, AOL, Baidu, Bing, DuckDuckGo, Kagi, Google, Yahoo!, Yandex D) LinkedIn, WhatsApp, Telegram
A) Use the same robots.txt for all subdomains B) Each subdomain must have its own robots.txt file C) Place a single robots.txt in the root directory D) Ignore robots.txt for subdomains
A) In the user's browser cache B) Inside each directory it applies to C) In the server's configuration files D) In the root of the web site hierarchy
A) LinkedIn, WhatsApp, Telegram B) Amazon, eBay, Alibaba C) Google, Facebook, Twitter D) Medium, Reddit, Yahoo
A) To enhance multimedia playback B) To prevent certain content from being misleading or irrelevant in search results C) To increase the number of visitors D) To improve server hardware
A) To indicate which portions of a website web crawlers are allowed to visit B) To display advertisements C) To enhance website security through encryption D) To store user login credentials
A) To manage which parts of a website are crawled and indexed B) To increase page load speed C) To encrypt data transmission D) To enhance visual design
A) RFC 9309 B) RFC 2616 C) RFC 7230 D) RFC 3986
A) The internet was small enough to maintain a complete list of all bots B) High bandwidth usage by video streaming C) Large file uploads by users D) Complex database queries
A) 2022 B) 2019 C) 1998 D) 2005
A) 1 megabyte B) 256 kilobytes C) 500 kibibytes (512000 bytes) D) Unlimited
A) JSON objects B) HTML tags C) A specific text-based format D) Binary code
A) All web pages are automatically indexed B) The website is blocked from search engines C) The server returns an error 404 D) Web robots assume that there are no limitations on crawling the entire site |