{"id":36,"date":"2004-08-12T17:19:11","date_gmt":"2004-08-13T00:19:11","guid":{"rendered":"http:\/\/redmonk.com\/sogrady\/wp\/?p=36"},"modified":"2004-08-12T17:19:11","modified_gmt":"2004-08-13T00:19:11","slug":"searching-audio","status":"publish","type":"post","link":"https:\/\/redmonk.com\/sogrady\/2004\/08\/12\/searching-audio\/","title":{"rendered":"Searching Audio"},"content":{"rendered":"<p>Jon Udell&#8217;s got a <a href=\"http:\/\/weblog.infoworld.com\/udell\/2004\/08\/12.html#a1058\">post<\/a> today that discusses the controversial <a href=\"http:\/\/lwn.net\/Articles\/95688\/\">Paul Graham<\/a> comments from <a href=\"http:\/\/conferences.oreillynet.com\/os2004\/\">OSCON<\/a> about Java. Jon&#8217;s point was mostly about the challenges of rich media with respect to searching and indexing. As he puts it:<\/p>\n<blockquote><p>the best way to pierce the opaqueness of large media files is to point into them, and then wrap the pointers with words that the search engines can see.<\/p><\/blockquote>\n<p>Can&#8217;t disagree with him there. But my basic problem with it is that it&#8217;s still a highly manual process. An individual has to decide what to link in <i>to<\/i>, and then decide what the keywords are for that particular comment. But unless the whole file is linked, I think things will get missed.<\/p>\n<p>As an example, I haven&#8217;t seen anybody yet discuss the comment from Dave Winer yet from his audio <a href=\"http:\/\/archive.scripting.com\/2004\/08\/08#When:7:00:48PM\">here<\/a> about 21:12 in, when he states that Microsoft perceives <a href=\"http:\/\/gmail.google.com\">Gmail<\/a> as a &#8220;serious threat.&#8221; And that &#8211; devoid of context &#8211; would be just another piece of speculation. But given where Winer <a href=\"http:\/\/blogs.law.harvard.edu\/crimson1\/2004\/08\/06#a2077\">was earlier in the week<\/a> it takes on a different level of significance. Without that context however, or a full index of the audio clip, things like that get missed.<\/p>\n<p>So Jon&#8217;s probably right in that this is the best way to make audio searchable right now, but it&#8217;s far from a perfect solution.<\/p>\n","protected":false},"excerpt":{"rendered":"<p>Jon Udell&#8217;s got a post today that discusses the controversial Paul Graham comments from OSCON about Java. Jon&#8217;s point was mostly about the challenges of rich media with respect to searching and indexing. As he puts it: the best way to pierce the opaqueness of large media files is to point into them, and then<\/p>\n","protected":false},"author":1,"featured_media":0,"comment_status":"open","ping_status":"closed","sticky":false,"template":"","format":"standard","meta":{"spay_email":"","footnotes":"","jetpack_publicize_message":"","jetpack_is_tweetstorm":false},"categories":[93],"tags":[],"class_list":["post-36","post","type-post","status-publish","format-standard","hentry","category-trends-observations"],"jetpack_featured_media_url":"","jetpack_publicize_connections":[],"jetpack_sharing_enabled":true,"_links":{"self":[{"href":"https:\/\/redmonk.com\/sogrady\/wp-json\/wp\/v2\/posts\/36","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/redmonk.com\/sogrady\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/redmonk.com\/sogrady\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/redmonk.com\/sogrady\/wp-json\/wp\/v2\/users\/1"}],"replies":[{"embeddable":true,"href":"https:\/\/redmonk.com\/sogrady\/wp-json\/wp\/v2\/comments?post=36"}],"version-history":[{"count":0,"href":"https:\/\/redmonk.com\/sogrady\/wp-json\/wp\/v2\/posts\/36\/revisions"}],"wp:attachment":[{"href":"https:\/\/redmonk.com\/sogrady\/wp-json\/wp\/v2\/media?parent=36"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/redmonk.com\/sogrady\/wp-json\/wp\/v2\/categories?post=36"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/redmonk.com\/sogrady\/wp-json\/wp\/v2\/tags?post=36"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}