Papers.
Research connected to its authors, projects, companies, talks, events, and the rest of the graph.
Add a paper ↗Semantics of query rewriting patterns in search logs
DOI 10.1145/2390148.2390153 · 4 citations · Source: openalex+authoritative-profileRicardo Baeza-Yates, Sumio Fujita, Georges Dupret, Ricardo Baeza‐Yates · 4 authors totalIDEAL
DOI 10.1145/2384916.2384955 · 25 citations · Source: openalex+authoritative-profileRicardo Baeza-Yates, Gaurang Kanvinde, Luz Rello, Ricardo Baeza‐Yates · 4 authors totalTargeted and scalable information dissemination in a distributed reputation mechanism
DOI 10.1145/2382536.2382547 · 10 citations · Source: openalex+orcid+dblp-identityJohan Pouwelse, Rahim Delaviz, Dick Epema · 3 authors totalBig data benchmarking
DOI 10.1145/2378356.2378368 · 13 citations · Source: openalex+career-authorityMilind Bhandarkar, Chaitan Baru, Raghunath Nambiar, Meikel Poess, Tilmann Rabl · 5 authors totalTowards energy-proportional datacenter memory with mobile DRAM
ISCA 2012 (ACM SIGARCH Computer Architecture News 40(3)) · DOI 10.1145/2366231.2337164 · 362 citations · Source: openalexFrank Austin Nothaft, Krishna T. Malladi, Karthika Periyathambi, Benjamin C. Lee, Christos Kozyrakis, Mark Horowitz · 6 authors totalA statistical similarity measure for aggregate crowd dynamics
ACM Transactions on Graphics · DOI 10.1145/2366145.2366209 · 126 citations · Source: openalexWe present an information-theoretic method to measure the similarity between a given set of observed, real-world data and visual simulation technique for aggregate crowd motions of a complex system consisting of many individual agents. This metric uses a two-step process to quantify a simulator's ability to reproduce the collective behaviors of the whole system, as observed in the recorded real-world data. First, Bayesian inference is used to estimate the simulation states which best correspond to the observed data, then a maximum likelihood estimator is used to approximate the prediction errors. This process is iterated using the EM-algorithm to produce a robust, statistical estimate of the magnitude of the prediction error as measured by its entropy (smaller is better). This metric serves as a simulator-to-data similarity measurement. We evaluated the metric in terms of robustness to sensor noise, consistency across different datasets and simulation methods, and correlation to perceptual metrics.
Jur van den Berg, Stephen J. Guy, Wenxi Liu, Rynson W. H. Lau, Ming C. Lin, Dinesh Manocha · 6 authors total(Big) usage data in web search
DOI 10.1145/2348283.2348531 · 2 citations · Source: openalex+authoritative-profileRicardo Baeza-Yates, Ricardo Baeza‐Yates, Yoelle Maarek · 3 authors totalSocial annotations: utility and prediction modeling
SIGIR · DOI 10.1145/2348283.2348324 · 25 citations · Source: dblp+semantic-scholarOmar Alonso, Patrick Pantel, Michael Gamon, Kevin Haas · 4 authors totalFinding trendsetters in information networks
DOI 10.1145/2339530.2339691 · 65 citations · Source: openalex+authoritative-profileRicardo Baeza-Yates, Diego Sáez-Trumper, Giovanni Comarela, Virgı́lio Almeida, Ricardo Baeza‐Yates, Fabrício Benevenuto · 6 authors totalRandom forests for metric learning with implicit pairwise position dependence
Knowledge Discovery and Data Mining · DOI 10.1145/2339530.2339680 · arXiv 1201.0610 · 74 citations · Source: arxiv+semantic-scholarMetric learning makes it plausible to learn semantically meaningful distances for complex distributions of data using label or pairwise constraint information. However, to date, most metric learning methods are based on a single Mahalanobis metric, which cannot handle heterogeneous data well. Those that learn multiple metrics throughout the feature space have demonstrated superior accuracy, but at a severe cost to computational efficiency. Here, we adopt a new angle on the metric learning problem and learn a single metric that is able to implicitly adapt its distance function throughout the feature space. This metric adaptation is accomplished by using a random forest-based classifier to underpin the distance function and incorporate both absolute pairwise position and standard relative position into the representation. We have implemented and tested our method against state of the art global and multi-metric methods on a variety of data sets. Overall, the proposed method outperforms both types of method in terms of accuracy (consistently ranked first) and is an order of magnitude faster than state of the art multi-metric methods (16x faster in the worst case).
Ran Xu, Caiming Xiong, David M. Johnson, Jason J. Corso · 4 authors totalExploring reflection in online communities
LAK · DOI 10.1145/2330601.2330630 · Source: dblp+adapt-autodesk-authorityAlex O'Connor, John McAuley, Alexander O'Connor, Dave Lewis 0001 · 4 authors totalLinked open corpus models, leveraging the semantic web for adaptive hypermedia
HT · DOI 10.1145/2309996.2310054 · Source: dblp+adapt-autodesk-authorityAlex O'Connor, Ian O'Keeffe, Alexander O'Connor, Philip Cass, Séamus Lawless, Vincent Wade · 6 authors totalEvaluation of a domain-aware approach to user model interoperability
HT · DOI 10.1145/2309996.2310030 · Source: dblp+adapt-autodesk-authorityAlex O'Connor, Eddie Walsh, Alexander O'Connor, Vincent Wade · 4 authors totalA trust-and-risk aware RBAC framework: tackling insider threat
ACM Symposium on Access Control Models and Technologies · DOI 10.1145/2295136.2295168 · 51 citations · Source: semantic-scholarNathalie Baracaldo, J. Joshi · 2 authors totalCAM: A Topology Aware Minimum Cost Flow Based Resource Manager for MapReduce Applications in the Cloud
ACM International Symposium on High Performance Distributed Computing · DOI 10.1145/2287076.2287110 · Source: acm+dblp+ibm-career-authorityDinesh Subhraveti, Min Li, Ali Raza Butt, Aleksandr Khasymski, Prasenjit Sarkar · 5 authors totalGPU accelerated AES-CBC for database applications.
SAC · DOI 10.1145/2245276.2245446 · Source: dblp+ubc-authorityRamon Lawrence, Scott Fazackerley, Steven M. McAvoy · 3 authors totalEntity matching for semistructured data in the Cloud
ACM SAC · DOI 10.1145/2245276.2245363 · 5 citations · Source: semantic-scholar+arxivSusan Malaika, M. Paradies, S. Malaika, Jérôme Siméon, S. Khatchadourian, K. Sattler · 6 authors totalFinding related tables
SIGMOD Conference · DOI 10.1145/2213836.2213962 · 217 citations · Source: semantic-scholarReynold Xin, A. Sarma, Lujun Fang, Nitin Gupta, A. Halevy, Hongrae Lee, Fei Wu, Cong Yu · 8 authors totalShark: fast data analysis using coarse-grained distributed memory
SIGMOD Conference · DOI 10.1145/2213836.2213934 · 141 citations · Source: semantic-scholarReynold Xin, Cliff Engle, Antonio Lupher, M. Zaharia, M. Franklin, S. Shenker, Ion Stoica · 7 authors totalGoogle's hybrid approach to research
Communications of the ACM · DOI 10.1145/2209249.2209262 · 59 citations · Source: semantic-scholar+openalexBy closely connecting research and development Google is able to conduct experiments on an unprecedented scale, often resulting in new capabilities for the company.
Peter Norvig, A. Spector, Slav Petrov · 3 authors totalLayout guidelines for web text and a web service to improve accessibility for dyslexics
DOI 10.1145/2207016.2207048 · 106 citations · Source: openalex+authoritative-profileRicardo Baeza-Yates, Luz Rello, Gaurang Kanvinde, Ricardo Baeza‐Yates · 4 authors totalOpen source column
ACM SIGMultimedia Records · DOI 10.1145/2206765.2206767 · 8 citations · Source: openalex+orcid+dblp-identityJohan Pouwelse, Niels Zeilemaker · 2 authors totalThe effect of links on networked user engagement
DOI 10.1145/2187980.2188167 · 7 citations · Source: openalex+authoritative-profileRicardo Baeza-Yates, Elad Yom‐Tov, Mounia Lalmas, Georges Dupret, Ricardo Baeza‐Yates, Pinard Donmez, Janette Lehmann · 7 authors totalLexical quality as a proxy for web text understandability
DOI 10.1145/2187980.2188142 · 22 citations · Source: openalex+authoritative-profileRicardo Baeza-Yates, Luz Rello, Ricardo Baeza‐Yates · 3 authors totalLeveraging trust and distrust for sybil-tolerant voting in online social media
DOI 10.1145/2185354.2185355 · 8 citations · Source: openalex+orcid+dblp-identityJohan Pouwelse, Nitin Chiluka, Nazareno Andrade, Henk Sips · 4 authors totalOn measuring the lexical quality of the web
DOI 10.1145/2184305.2184307 · 16 citations · Source: openalex+authoritative-profileRicardo Baeza-Yates, Ricardo Baeza‐Yates, Luz Rello · 3 authors totalIdentification of top relevant temporal expressions in documents
TempWeb · DOI 10.1145/2169095.2169102 · 39 citations · Source: dblp+semantic-scholarOmar Alonso, Jannik Strötgen, Michael Gertz · 3 authors totalYour mouse is a database
Communications of the ACM · DOI 10.1145/2160718.2160735 · 40 citations · Source: openalex+semantic-scholarWeb and mobile applications are increasingly composed of asynchronous and real-time streaming services and push notifications.
Erik Meijer · 1 author totalIdempotence Is Not a Medical Condition
ACM Queue · DOI 10.1145/2160718.2160734 · 66 citations · Source: semantic-scholarExplains why at-least-once message delivery plus idempotent processing is the practical foundation of reliable distributed applications.
Pat Helland · 1 author totalShort message communications
DOI 10.1145/2160601.2160607 · 23 citations · Source: openalex+personal-publication-listRob Munro, Robert Munro, Christopher D. Manning · 3 authors totalConcurrent tries with efficient non-blocking snapshots
DOI 10.1145/2145816.2145836 · 109 citations · Source: openalexWe describe a non-blocking concurrent hash trie based on shared-memory single-word compare-and-swap instructions. The hash trie supports standard mutable lock-free operations such as insertion, removal, lookup and their conditional variants. To ensure space-efficiency, removal operations compress the trie when necessary.
Martin Odersky, Aleksandar Prokopec, Nathan Bronson, Phil Bagwell · 4 authors totalScala-virtualized
PEPM 2012 (Partial Evaluation and Program Manipulation) · DOI 10.1145/2103746.2103769 · 55 citations · Source: openalexScala-Virtualized extends the Scala language to better support hosting embedded DSLs. Embedding a DSL in Scala-Virtualized comes with all the benefits of a shallow embedding thanks to Scala's flexible syntax, without giving up analyzing and manipulating the domain program -- typically exclusive to deep embeddings. Through lightweight modular staging, implemented in standard Scala, the benefits of a deep embedding are recovered with little overhead. Scala-Virtualized lifts more of the language's built-in constructs and static information to complete this support and make it more convenient. We illustrate how Scala-Virtualized makes Scala an even better host for embedded DSLs along three axes of customizing the language: syntax, run-time behavior and static semantics.
Adriaan Moors, Tiark Rompf, Philipp Haller, Martin Odersky · 4 authors totalStagedSAC: a case study in performance-oriented DSL development
ACM SIGPLAN Workshop on Partial Evaluation and Program Manipulation · DOI 10.1145/2103746.2103762 · 14 citations · Source: semantic-scholarVlad Ureche, Tiark Rompf, Arvind K. Sujeeth, Hassan Chafi, Martin Odersky · 5 authors totalTop-10 Data Mining Case Studies
International Journal of Information Technology and Decision Making · DOI 10.1142/S021962201240007X · 6 citations · Source: personal-publication-catalog+semantic-scholarGabor Melli, Xindong Wu, P. Beinat, F. Bonchi, Longbing Cao, Rong Duan, C. Faloutsos, R. Ghani · 14 authors totalRoles of Performance and Human Capital in College Football Coaches' Compensation.
DOI 10.1123/JSM.27.1.73 · 27 citations · Source: semantic-scholarDespite the escalation of football coaches’ salaries at National Collegiate Athletic Association (NCAA) Football Bowl Subdivision (FBS) institutions, little empirical investigation has been undertaken to identify the determinants of their compensation. As such, the purpose of this study is to explain how the level of coaching compensation is determined based on three theoretical perspectives in managerial compensation: marginal productivity theory, human capital theory, and managerialism. The analysis of compensation data of head football coaches at FBS institutions in 2006–2007 shows that the maximum total compensation of these coaches increases with their past performance. The results further reveal that coaches with greater human capital tend to receive a compensation package where bonuses account for a smaller proportion of the maximum total compensation. Overall, these findings mostly confirm the predictions drawn from managerial productivity theory, human capital theory and managerialism.
Jose Plehn, Yuhei Inoue, J. Plehn-Dujowich, A. Kent, Steve Swanson · 5 authors totalThe use of novel, direct diode lasers for large area hard-facing and high deposition rate cladding to enhance surface wear and corrosion resistance
SPIE LASE (Proc. SPIE 8239, High-Power Diode Laser Technology and Applications X) · DOI 10.1117/12.908947 · 11 citations · Source: semantic-scholar+crossrefWolfgang Juchmann, Stephen Brookshier, John Washko, Keith Parker, Frank Gaebler · 5 authors totalColorless green ideas learn furiously: Chomsky and the two cultures of statistical learning
DOI 10.1111/j.1740-9713.2012.00590.x · 19 citations · Source: semantic-scholarPeter Norvig · 1 author totalMapping XML to a Wide Sparse Table
IEEE Transactions on Knowledge and Data Engineering · DOI 10.1109/tkde.2012.221 · 5 citations · Source: openalex+career-authorityNikita Shamgunov, Liang Jeff Chen, Philip A. Bernstein, Peter Carlin, Dimitrije Filipovic, Michael Rys, James F. Terwilliger, Milos Todic · 10 authors totalScalable Learning of Collective Behavior
IEEE Trans. Knowl. Data Eng. · DOI 10.1109/TKDE.2011.38 · Source: dblp+asu-first-party+career-authorityLei Tang, Xufei Wang, Huan Liu · 3 authors totalIdentifying Evolving Groups in Dynamic Multimode Networks
IEEE Trans. Knowl. Data Eng. · DOI 10.1109/TKDE.2011.159 · Source: dblp+asu-first-party+career-authorityLei Tang, Huan Liu, Jianping Zhang · 3 authors totalUser Taglines: Alternative Presentations of Expertise and Interest in Social Media
SocialInformatics · DOI 10.1109/SocialInformatics.2012.68 · arXiv 1212.1927 · 14 citations · Source: dblp+semantic-scholarWeb applications are increasingly showing recommended users from social media along with some descriptions, an attempt to show relevancy-why they are being shown. For example, Twitter search for a topical keyword shows expert twitterers on the side for `whom to follow'. Google+ and Facebook also recommend users to follow or add to friend circle. Popular Internet newspaper-The Huffing ton Post shows Twitter experts on the side of an article for authoritative relevant tweets. The state of the art shows user profile bio as summary for Twitter experts, but it has issues with length constraints imposed by the user interface (UI) design, missing bio and sometimes funny profile bio. Alternatively, applications can use human generated user summary, but it will not scale. Therefore, we study the problem of automatic generation of informative expertise summary or taglines for Twitter experts in space constraint imposed by UI design. We propose three methods for expertise summary generation: Occupation-Pattern based, Link-Triangulation based and User-Classification based, with the use of knowledge-enhanced computing approaches. We also propose methods for final summary selection for users with multiple candidates of generated summaries and evaluate results by user-study for both generation and selection tasks. The results of proposed tagline generation methods show 92.8% good summaries with majority agreement in the best case and 70% in the worst case while outperforming the state of the art up to 88%. This study has implications in the area of expert profiling, user presentation and application design for engaging user experience.
Omar Alonso, Hemant Purohit, A. Dow, Lei Duan, Kevin Haas · 5 authors totalDo You Know the Way to SNA?: A Process Model for Analyzing and Visualizing Social Media Network Data
DOI 10.1109/socialinformatics.2012.26 · 69 citations · Source: openalex+first-party-career-authorityMarc Smith, Derek L. Hansen, Dana Rotman, Elizabeth Bonsignore, N. Milić-Frayling, Eduarda Mendes Rodrigues, Marc A. Smith, Ben Shneiderman · 8 authors totalPerformance analysis of the Libswift P2P streaming protocol
DOI 10.1109/p2p.2012.6335790 · 18 citations · Source: openalex+orcid+dblp-identityJohan Pouwelse, Riccardo Petrocco, Dick Epema · 3 authors totalNeuFlow: Dataflow vision processing system-on-a-chip
IEEE International Midwest Symposium on Circuits and Systems (MWSCAS) · DOI 10.1109/MWSCAS.2012.6292202 · 94 citations · Source: dblpClément Farabet, Phi-Hung Pham, Darko Jelaca, Berin Martini, Yann LeCun, Eugenio Culurciello · 6 authors totalSensing the "Health State" of a Community
IEEE pervasive computing · DOI 10.1109/MPRV.2011.79 · 258 citations · Source: semantic-scholar+dblp+career-authoritySai Moturu, Anmol Madan, Manuel Cebrian, S. Moturu, K. Farrahi, A. Pentland · 6 authors totalPlay2: A New Era of Web Application Development
IEEE Internet Computing · DOI 10.1109/MIC.2012.84 · Source: ieee+author-copy+play-frameworkSadek Aldrobi, Sadek Drobi · 2 authors totalA Visual Tool for Querying and Exploring XML Data
DOI 10.1109/la-web.2012.20 · 1 citations · Source: openalex+authoritative-profileRicardo Baeza-Yates, Ricardo Baeza‐Yates, C. Baldenegro Barrera, Valeria Herskovic · 4 authors totalGraphGen: A Tool for Automatic Generation of Multipartite Graphs from Arbitrary Data
DOI 10.1109/la-web.2012.15 · 2 citations · Source: openalex+authoritative-profileRicardo Baeza-Yates, Sandra Álvarez-García, Ricardo Baeza‐Yates, Nieves R. Brisaboa, Josep-L. Larriba-Pey, Óscar Pedreira · 6 authors totalExploiting Distributional Semantic Models in Question Answering
DOI 10.1109/icsc.2012.53 · 11 citations · Source: openalexThis paper investigates the role of Distributional Semantic Models (DSMs) in Question Answering (QA), and specifically in a QA system called Question Cube. Question Cube is a framework for QA that combines several techniques to retrieve passages containing the exact answers for natural language questions. It exploits Information Retrieval models to seek candidate answers and Natural Language Processing algorithms for the analysis of questions and candidate answers both in English and Italian. The data source for the answer is an unstructured text document collection stored in search indices. In this paper we propose to exploit DSMs in the Question Cube framework. In DSMs words are represented as mathematical points in a geometric space, also known as semantic space. Words are similar if they are close in that space. Our idea is that DSMs approaches can help to compute relatedness between users' questions and candidate answers by exploiting paradigmatic relations between words. Results of an experimental evaluation carried out on CLEF2010 QA dataset, prove the effectiveness of the proposed approach.
Piero Molino, Pierpaolo Basile, Annalina Caputo, Pasquale Lops, Giovanni Semeraro · 5 authors totalEstimating probability of collision for safe motion planning under Gaussian motion and sensing uncertainty
ICRA 2012 · DOI 10.1109/icra.2012.6224727 · 115 citations · Source: openalexWe present a fast, analytical method for estimating the probability of collision of a motion plan for a mobile robot operating under the assumptions of Gaussian motion and sensing uncertainty. Estimating the probability of collision is an integral step in many algorithms for motion planning under uncertainty and is crucial for characterizing the safety of motion plans. Our method is computationally fast, enabling its use in online motion planning, and provides conservative estimates to promote safety. To improve accuracy, we use a novel method to truncate estimated a priori state distributions to account for the fact that the probability of collision at each stage along a plan is conditioned on the previous stages being collision free. Our method can be directly applied within a variety of existing motion planners to improve their performance and the quality of computed plans. We apply our method to a car-like mobile robot with second order dynamics and to a steerable medical needle in 3D and demonstrate that our method for estimating the probability of collision is orders of magnitude faster than naïve Monte Carlo sampling methods and reduces estimation error by more than 25% compared to prior methods.
Jur van den Berg, Sachin Patil, Ron Alterovitz · 3 authors total