{"id":{"repo_id":"uiuc","oai_identifier":"oai:www.ideals.illinois.edu:2142/98173"},"canonical_url":"https://search.dev.ndltd.org/etd/uiuc/oai:www.ideals.illinois.edu:2142/98173","repository":{"repo_id":"uiuc","name":"University of Illinois - Urbana-Champaign","base_url":"https://www.ideals.illinois.edu/oai-pmh"},"display":{"title":"Discovering hidden patterns in technology management phenomena: three essays on using data analytics for exploration of causal inferences","abstract":"The availability of copious amounts of data, increased computational power to analyze them, and the readiness of various techniques to extract information and find patterns among them, are having a significant impact on the way companies are managed. However, the ability of these methods to generate causal inferences is still one of the fundamental questions that must be resolved before its mainstream incorporation into quantitative research in management. To answer the question -- of what represents the current data deluge and data analytics methods for management research from both a philosophical and methodological perspective -- is the main purpose of this dissertation. In order to address that question, from a philosophical perspective, we adopt a realist view of causation. From a methodological perspective, we argue that the use of data analytics methods should start with a phase strongly grounded in theory and end with the abduction of new theories from the patterns. In the three essays of this dissertation, we exemplify the use of data analytics methods in the study of technology management phenomena. The first essay focuses on the phenomenon of open source software development. Despite the tremendous popularity of certain open source software applications, one major challenge for Free/Libre Open Source Software (FLOSS) online communities has been the high mortality rate of initiatives, as observed in the decline and sometimes fading away of development activity for the clear majority of projects over time. Another challenge is that the sheer number of simultaneously coexisting projects in these communities highlights the importance of visibility as a way for a project to distinguish itself from the collection. In this study, we examine the relationship between endurance and visibility connected to keyword characteristics. Our results suggest that higher interest of users is positively associated with the endurance of the projects in the communities. Further, we find that the selection of keywords, reflecting the functionality of the software and the operating system, is strongly associated with user attention. Per our study, increasing the visibility of the project is an important mechanism to sustain its activity over time. The second essay centers on exclusivity in technology licensing. While prior research has significantly advanced our understanding about exclusivity in licensing, there are still significant gaps in our knowledge about how licensing exclusivity is impacted by the interplay between different contextual and intrinsic attributes of the license. Exclusivity in licensing can be highly complex and contingent, potentially reflecting the interactions between different theoretical explanations, and the boundary conditions that apply to each theory. The exploration of such contingencies and complexities is hampered in conventional econometric analyses, which we seek to overcome by employing a novel empirical technique called decision tree induction, a powerful machine learning tool for uncovering nested “multiple theoretical viewpoints.” Implications for the empirical and theoretical literature on licensing, and for abductive theory development by leveraging “big data” are discussed. The third and final essay addresses alliance formation in the computer services industry. Strategic alliances have steadily increased over the last three decades as popular instruments for interfirm cooperation. While there are competing and complementary theoretical bodies of work that have addressed this phenomenon, the question of who allies with whom is still relevant due to the complexities surrounding the phenomenon and the challenging nature of the prediction task for alliance formation. Social network approaches have substantively contributed to our understanding of the partner choice in alliances; however, they also bring forth some limitations. We extend that previous work by addressing some of them through the introduction of the concept of heterogeneous networks and the application of a novel machine learning intensive technique to predict alliance formation. Our results suggest a high predictive accuracy of the technique. Implications for the path dependence of alliance formation processes are also discussed.","abstract_html":"The availability of copious amounts of data, increased computational power to analyze them, and the readiness of various techniques to extract information and find patterns among them, are having a significant impact on the way companies are managed. However, the ability of these methods to generate causal inferences is still one of the fundamental questions that must be resolved before its mainstream incorporation into quantitative research in management. To answer the question -- of what represents the current data deluge and data analytics methods for management research from both a philosophical and methodological perspective -- is the main purpose of this dissertation. In order to address that question, from a philosophical perspective, we adopt a realist view of causation. From a methodological perspective, we argue that the use of data analytics methods should start with a phase strongly grounded in theory and end with the abduction of new theories from the patterns. In the three essays of this dissertation, we exemplify the use of data analytics methods in the study of technology management phenomena. The first essay focuses on the phenomenon of open source software development. Despite the tremendous popularity of certain open source software applications, one major challenge for Free/Libre Open Source Software (FLOSS) online communities has been the high mortality rate of initiatives, as observed in the decline and sometimes fading away of development activity for the clear majority of projects over time. Another challenge is that the sheer number of simultaneously coexisting projects in these communities highlights the importance of visibility as a way for a project to distinguish itself from the collection. In this study, we examine the relationship between endurance and visibility connected to keyword characteristics. Our results suggest that higher interest of users is positively associated with the endurance of the projects in the communities. Further, we find that the selection of keywords, reflecting the functionality of the software and the operating system, is strongly associated with user attention. Per our study, increasing the visibility of the project is an important mechanism to sustain its activity over time. The second essay centers on exclusivity in technology licensing. While prior research has significantly advanced our understanding about exclusivity in licensing, there are still significant gaps in our knowledge about how licensing exclusivity is impacted by the interplay between different contextual and intrinsic attributes of the license. Exclusivity in licensing can be highly complex and contingent, potentially reflecting the interactions between different theoretical explanations, and the boundary conditions that apply to each theory. The exploration of such contingencies and complexities is hampered in conventional econometric analyses, which we seek to overcome by employing a novel empirical technique called decision tree induction, a powerful machine learning tool for uncovering nested “multiple theoretical viewpoints.” Implications for the empirical and theoretical literature on licensing, and for abductive theory development by leveraging “big data” are discussed. The third and final essay addresses alliance formation in the computer services industry. Strategic alliances have steadily increased over the last three decades as popular instruments for interfirm cooperation. While there are competing and complementary theoretical bodies of work that have addressed this phenomenon, the question of who allies with whom is still relevant due to the complexities surrounding the phenomenon and the challenging nature of the prediction task for alliance formation. Social network approaches have substantively contributed to our understanding of the partner choice in alliances; however, they also bring forth some limitations. We extend that previous work by addressing some of them through the introduction of the concept of heterogeneous networks and the application of a novel machine learning intensive technique to predict alliance formation. Our results suggest a high predictive accuracy of the technique. Implications for the path dependence of alliance formation processes are also discussed.","abstract_has_math":false,"creators":["Fernandez Corrales, Carla Beatriz"],"institution":"University of Illinois at Urbana-Champaign","degree_name":"Ph.D.","degree_level":"Dissertation","degree_discipline":"Business Administration","degree_department":null,"school":null,"contributors":["Subramanyam, Ramanath","Somaya, Deepak","Larson, Eric C.","Mahoney, Joseph T."],"advisors":[],"committee_chairs":[],"committee_members":[],"year":2017,"date_issued":"2017-09-29T17:45:13Z","date_published":"2017-09-29T17:45:13Z","updated_at":"2026-07-22T22:24:35Z","subjects":["Data analytics","Machine learning","Causal inferences","Abduction of theories","Technology management","Alliance formation","Technology licencing","Open source software"],"languages":["en"],"rights":["Copyright 2017 Carla Beatriz Fernández Corrales"],"rights_urls":[],"identifier_entries":[]},"links":{"outbound_url":"http://hdl.handle.net/2142/98173","outbound_label":"Handle","outbound_source":"dc:identifier"},"metadata_groups":[{"id":"people","label":"People","entries":[{"key":"dc:contributor","label":"Contributor","values":["Subramanyam, Ramanath","Somaya, Deepak","Larson, Eric C.","Mahoney, Joseph T."]},{"key":"dc:creator","label":"Author","values":["Fernandez Corrales, Carla Beatriz"]}]},{"id":"academic_context","label":"Academic Context","entries":[{"key":"dc:date","label":"Dc Date","values":["2017-09-29T17:45:13Z","2020-03-03T10:15:11Z","2017-06-28","2017-08"]},{"key":"dc:type","label":"Dc Type","values":["text"]},{"key":"thesis:degree_discipline","label":"Discipline","values":["Business Administration"]},{"key":"thesis:degree_level","label":"Degree Level","values":["Dissertation"]},{"key":"thesis:degree_name","label":"Degree Name","values":["Ph.D."]},{"key":"thesis:institution_name","label":"Thesis Institution Name","values":["University of Illinois at Urbana-Champaign"]}]},{"id":"subjects_keywords","label":"Subjects and Keywords","entries":[{"key":"dc:subject","label":"Dc Subject","values":["Data analytics","Machine learning","Causal inferences","Abduction of theories","Technology management","Alliance formation","Technology licencing","Open source software"]}]},{"id":"language_rights","label":"Language and Rights","entries":[{"key":"dc:language","label":"Dc Language","values":["en"]},{"key":"dc:rights","label":"Dc Rights","values":["Copyright 2017 Carla Beatriz Fernández Corrales"]}]},{"id":"identifiers","label":"Identifiers","entries":[{"key":"dc:identifier","label":"Identifier","values":["http://hdl.handle.net/2142/98173"]}]},{"id":"additional","label":"Additional Metadata","entries":[{"key":"dc:description","label":"Description","values":["The availability of copious amounts of data, increased computational power to analyze them, and the readiness of various techniques to extract information and find patterns among them, are having a significant impact on the way companies are managed. However, the ability of these methods to generate causal inferences is still one of the fundamental questions that must be resolved before its mainstream incorporation into quantitative research in management. To answer the question -- of what represents the current data deluge and data analytics methods for management research from both a philosophical and methodological perspective -- is the main purpose of this dissertation. In order to address that question, from a philosophical perspective, we adopt a realist view of causation. From a methodological perspective, we argue that the use of data analytics methods should start with a phase strongly grounded in theory and end with the abduction of new theories from the patterns. In the three essays of this dissertation, we exemplify the use of data analytics methods in the study of technology management phenomena. The first essay focuses on the phenomenon of open source software development. Despite the tremendous popularity of certain open source software applications, one major challenge for Free/Libre Open Source Software (FLOSS) online communities has been the high mortality rate of initiatives, as observed in the decline and sometimes fading away of development activity for the clear majority of projects over time. Another challenge is that the sheer number of simultaneously coexisting projects in these communities highlights the importance of visibility as a way for a project to distinguish itself from the collection. In this study, we examine the relationship between endurance and visibility connected to keyword characteristics. Our results suggest that higher interest of users is positively associated with the endurance of the projects in the communities. Further, we find that the selection of keywords, reflecting the functionality of the software and the operating system, is strongly associated with user attention. Per our study, increasing the visibility of the project is an important mechanism to sustain its activity over time. The second essay centers on exclusivity in technology licensing. While prior research has significantly advanced our understanding about exclusivity in licensing, there are still significant gaps in our knowledge about how licensing exclusivity is impacted by the interplay between different contextual and intrinsic attributes of the license. Exclusivity in licensing can be highly complex and contingent, potentially reflecting the interactions between different theoretical explanations, and the boundary conditions that apply to each theory. The exploration of such contingencies and complexities is hampered in conventional econometric analyses, which we seek to overcome by employing a novel empirical technique called decision tree induction, a powerful machine learning tool for uncovering nested “multiple theoretical viewpoints.” Implications for the empirical and theoretical literature on licensing, and for abductive theory development by leveraging “big data” are discussed. The third and final essay addresses alliance formation in the computer services industry. Strategic alliances have steadily increased over the last three decades as popular instruments for interfirm cooperation. While there are competing and complementary theoretical bodies of work that have addressed this phenomenon, the question of who allies with whom is still relevant due to the complexities surrounding the phenomenon and the challenging nature of the prediction task for alliance formation. Social network approaches have substantively contributed to our understanding of the partner choice in alliances; however, they also bring forth some limitations. We extend that previous work by addressing some of them through the introduction of the concept of heterogeneous networks and the application of a novel machine learning intensive technique to predict alliance formation. Our results suggest a high predictive accuracy of the technique. Implications for the path dependence of alliance formation processes are also discussed.","Submission published under a 24 month embargo labeled 'U of I Access', the embargo will last until 2019-08-01","The student, Carla Fernandez Corrales, accepted the attached license on 2017-06-22 at 17:37.","The student, Carla Fernandez Corrales, submitted this Dissertation for approval on 2017-06-22 at 17:47.","This Dissertation was approved for publication on 2017-06-28 at 13:16.","DSpace SAF Submission Ingestion Package generated from Vireo submission #11235 on 2017-09-29 at 10:45:56","Made available in DSpace on 2017-09-29T17:45:13Z (GMT). No. of bitstreams: 3 FERNANDEZCORRALES-DISSERTATION-2017.pdf: 2765348 bytes, checksum: f951b563066a6f9c111f25428f6d4e21 (MD5) LICENSE.txt: 4221 bytes, checksum: 1981382f4db3750ea6fe03b7f0da2bd9 (MD5) PROQUEST_LICENSE.txt: 4567 bytes, checksum: fd86993ec48e88aad57528617f72bb69 (MD5) Previous issue date: 2017-06-28","Embargo set by: Colleen Fallaw for item 103446 Lift date: 2019-09-29T17:48:06Z Reason: Author requested U of Illinois access only (OA after 2yrs) in Vireo ETD system","Embargo set by: Seth Robbins for item 103446 Lift date: 2020-03-02T19:56:41Z Reason: Author requested U of Illinois access only (OA after 2yrs) in Vireo ETD system","Embargo set by: Seth Robbins for item 103446 Lift date: 2020-03-02T19:59:52Z Reason: Author requested U of Illinois access only (OA after 2yrs) in Vireo ETD system","Embargo set by: Seth Robbins for item 103446 Lift date: 2020-03-02T20:02:46Z Reason: Author requested U of Illinois access only (OA after 2yrs) in Vireo ETD system","U of I Only Restriction Lifted for Item 103446 on 2020-03-03T10:15:11Z."]},{"key":"dc:format","label":"Dc Format","values":["application/pdf"]},{"key":"dc:title","label":"Title","values":["Discovering hidden patterns in technology management phenomena: three essays on using data analytics for exploration of causal inferences"]}]}],"canonical_facts":{"dc:contributor":["Subramanyam, Ramanath","Somaya, Deepak","Larson, Eric C.","Mahoney, Joseph T."],"dc:creator":["Fernandez Corrales, Carla Beatriz"],"dc:date":["2017-09-29T17:45:13Z","2020-03-03T10:15:11Z","2017-06-28","2017-08"],"dc:description":["The availability of copious amounts of data, increased computational power to analyze them, and the readiness of various techniques to extract information and find patterns among them, are having a significant impact on the way companies are managed. However, the ability of these methods to generate causal inferences is still one of the fundamental questions that must be resolved before its mainstream incorporation into quantitative research in management. To answer the question -- of what represents the current data deluge and data analytics methods for management research from both a philosophical and methodological perspective -- is the main purpose of this dissertation. In order to address that question, from a philosophical perspective, we adopt a realist view of causation. From a methodological perspective, we argue that the use of data analytics methods should start with a phase strongly grounded in theory and end with the abduction of new theories from the patterns. In the three essays of this dissertation, we exemplify the use of data analytics methods in the study of technology management phenomena. The first essay focuses on the phenomenon of open source software development. Despite the tremendous popularity of certain open source software applications, one major challenge for Free/Libre Open Source Software (FLOSS) online communities has been the high mortality rate of initiatives, as observed in the decline and sometimes fading away of development activity for the clear majority of projects over time. Another challenge is that the sheer number of simultaneously coexisting projects in these communities highlights the importance of visibility as a way for a project to distinguish itself from the collection. In this study, we examine the relationship between endurance and visibility connected to keyword characteristics. Our results suggest that higher interest of users is positively associated with the endurance of the projects in the communities. Further, we find that the selection of keywords, reflecting the functionality of the software and the operating system, is strongly associated with user attention. Per our study, increasing the visibility of the project is an important mechanism to sustain its activity over time. The second essay centers on exclusivity in technology licensing. While prior research has significantly advanced our understanding about exclusivity in licensing, there are still significant gaps in our knowledge about how licensing exclusivity is impacted by the interplay between different contextual and intrinsic attributes of the license. Exclusivity in licensing can be highly complex and contingent, potentially reflecting the interactions between different theoretical explanations, and the boundary conditions that apply to each theory. The exploration of such contingencies and complexities is hampered in conventional econometric analyses, which we seek to overcome by employing a novel empirical technique called decision tree induction, a powerful machine learning tool for uncovering nested “multiple theoretical viewpoints.” Implications for the empirical and theoretical literature on licensing, and for abductive theory development by leveraging “big data” are discussed. The third and final essay addresses alliance formation in the computer services industry. Strategic alliances have steadily increased over the last three decades as popular instruments for interfirm cooperation. While there are competing and complementary theoretical bodies of work that have addressed this phenomenon, the question of who allies with whom is still relevant due to the complexities surrounding the phenomenon and the challenging nature of the prediction task for alliance formation. Social network approaches have substantively contributed to our understanding of the partner choice in alliances; however, they also bring forth some limitations. We extend that previous work by addressing some of them through the introduction of the concept of heterogeneous networks and the application of a novel machine learning intensive technique to predict alliance formation. Our results suggest a high predictive accuracy of the technique. Implications for the path dependence of alliance formation processes are also discussed.","Submission published under a 24 month embargo labeled 'U of I Access', the embargo will last until 2019-08-01","The student, Carla Fernandez Corrales, accepted the attached license on 2017-06-22 at 17:37.","The student, Carla Fernandez Corrales, submitted this Dissertation for approval on 2017-06-22 at 17:47.","This Dissertation was approved for publication on 2017-06-28 at 13:16.","DSpace SAF Submission Ingestion Package generated from Vireo submission #11235 on 2017-09-29 at 10:45:56","Made available in DSpace on 2017-09-29T17:45:13Z (GMT). No. of bitstreams: 3 FERNANDEZCORRALES-DISSERTATION-2017.pdf: 2765348 bytes, checksum: f951b563066a6f9c111f25428f6d4e21 (MD5) LICENSE.txt: 4221 bytes, checksum: 1981382f4db3750ea6fe03b7f0da2bd9 (MD5) PROQUEST_LICENSE.txt: 4567 bytes, checksum: fd86993ec48e88aad57528617f72bb69 (MD5) Previous issue date: 2017-06-28","Embargo set by: Colleen Fallaw for item 103446 Lift date: 2019-09-29T17:48:06Z Reason: Author requested U of Illinois access only (OA after 2yrs) in Vireo ETD system","Embargo set by: Seth Robbins for item 103446 Lift date: 2020-03-02T19:56:41Z Reason: Author requested U of Illinois access only (OA after 2yrs) in Vireo ETD system","Embargo set by: Seth Robbins for item 103446 Lift date: 2020-03-02T19:59:52Z Reason: Author requested U of Illinois access only (OA after 2yrs) in Vireo ETD system","Embargo set by: Seth Robbins for item 103446 Lift date: 2020-03-02T20:02:46Z Reason: Author requested U of Illinois access only (OA after 2yrs) in Vireo ETD system","U of I Only Restriction Lifted for Item 103446 on 2020-03-03T10:15:11Z."],"dc:format":["application/pdf"],"dc:identifier":["http://hdl.handle.net/2142/98173"],"dc:language":["en"],"dc:rights":["Copyright 2017 Carla Beatriz Fernández Corrales"],"dc:subject":["Data analytics","Machine learning","Causal inferences","Abduction of theories","Technology management","Alliance formation","Technology licencing","Open source software"],"dc:title":["Discovering hidden patterns in technology management phenomena: three essays on using data analytics for exploration of causal inferences"],"dc:type":["text"],"thesis:degree_discipline":["Business Administration"],"thesis:degree_level":["Dissertation"],"thesis:degree_name":["Ph.D."],"thesis:institution_name":["University of Illinois at Urbana-Champaign"]},"updated_at":"2026-07-22T22:24:35Z"}