1,720,987 research outputs found
Data Mining in Traffic Accident-A Case Study for Bus
臺灣地區道路交通工程不斷改進,運具方面的結構也有也改,然而運輸問題卻日益嚴重。根據警政署的統計資料,每年交通事故不斷的攀昇,平均約2800人死於交通事故,而大客車在每萬輛死肇肇事率都遠高於其他車種,所以一旦發生事故,其所造成的傷亡往往遠多於其他車種。 大客車事故頻傳的因素不外乎就是超速、酒醉駕車、違規駕駛、剎車失靈等,因此本研究嘗試資料探勘於交通事故之分析,探討大客車事故發生的主要因素。利用群集分析找出同質性最高的肇事集合,以此結果為基礎分析,並用卡方檢定驗證群集的正確性;然後套入判別分析中,作出其判別函數及預測其分類正確率,並找出影響大客車的肇事變數。 本研究採用警政署自民國92年至96年共六年全國大客車交通事故資料,資料總件數為15514件。研究結果為92至96年的訓練樣本和測試樣本的分類正確率都達90%以上,代表其判別率佳;影響六年的肇事變數為肇事原因、發生月份、道路類別、速限、分向設施為主要因素,其中,又以肇事原因佔的比例最高。在肇事原因中,主要發生的原因為變換車道或方向不當、未保持行車安全距踓、酒醉後駕駛失控、違反號誌管制或指揮、違反特定標誌(線)標制、其他駕駛人因素、非駕駛人因素,並針對此研究結果提出改善策略。According to the statistic data from National Police Agency, traffic accidents increase year over year. Averagely, 2,800 people die in traffic accidents. Moreover, in every thousand cars that cause deadly car accidents, bus is the main trouble maker. Thus, once a car accident happens, bus accidents usually lead to more deaths than other car accidents. Bus accidents usually include speeding, drunk driving, driving against traffic regulations, and brake failing…etc. Therefore, in this thesis, I try to do traffic accidents analysis by using data mining, and figure out the main causes of bus accidents. In addition, I use cluster analysis to find out the most homogeneous class which causes bus accidents. As a result, I use chi-square test to verify the validity of the cluster and apply it to discriminant analysis which can determine the discriminant function and predict the accuracy of classification, so that the variable which causes bus accidents will come out. In this research, the bus accidents data from 2003 to 2007 are included; the total number of the data is 15514, which can be found in National Police Agency. The result of this research includes the Training Samples and the Test Samples from 2003 to 2007, which’s accuracy of classification are all over ninety percent. Therefore, the identification is quite excellent. The major variables which cause bus accidents from 2003 to 2007 are accident causes, which months, what kind of road, speed limit, and curb systems. Among these factors, accident causes are the most. And results for this strategy to improve.誌謝 I要 IIIbstract IV一章 緒論 1.1 研究背景與動機 1.2 研究目的 3.3 研究範圍 3.4 研究內容與流程 3二章 文獻回顧 7.1 資料探勘相關文獻 7.1.1 資料探勘的定義 7.1.2 資料探勘的步驟與資料庫知識發現 9.1.3 資料探勘的功能與技術 10.1.4 資料探勘的應用 13.2 交通事故相關文獻 16.2.1 交通事故定義及分類 16.2.2 肇事分析之應用 19.2.3 國內道路遊覽車肇事案件分析 21.3 綜合評析 25三章 研究方法 29.1 群集分析 29.2 判別分析 39.3 因子分析 43.4小結 48四章 肇事資料前置處理 49.1 分析流程 49.2 肇事資料蒐集 50.3 肇事資料前置處理 51.4 肇事資料轉換 52.5 肇事變數 54五章 資料分析 57.1 群集方法與分析流程 57.2 群集數目選擇 58.3 群集結果 61.4 群集驗證 68.5 判別模式架構與分析流程 76.6 判別函數的構建 77.7 判別分析結果 78.8 判別模式綜合分析 93六章 案例分析 103.1臺灣地區 103.2臺北市 108.3南投縣 114.4案例比較 119七章 結論與建議 121.1結論 121.2 建議 122考文獻 12
Going Beyond Counting First Authors in Author Co-citation Analysis
The present study examines one of the fundamental aspects of author co-citation analysis (ACA) - the way co-citation
counts are defined. Co-citation counting provides the data on which all subsequent statistical analyses and mappings
are based, and we compare ACA results based on two different types of co-citation counting - the traditional type that
only counts the first one among a cited work's authors on the one hand and a non-traditional type that takes into
account the first 5 authors of a cited work on the other hand. Results indicate that the picture produced through this non-traditional author co-citation counting contains more coherent author groups and is therefore considerably clearer. However, this picture represents fewer specialties in the research field being studied than that produced through the traditional first-author co-citation counting when the same number of top-ranked authors is selected and analyzed. Reasons for these effects are discussed
Variations on the Author
“Variations on the Author” discusses two of Eduardo Coutinho’s recent films (Um Dia na Vida, from 2010, and Últimas Conversas, posthumously released in 2015) and their contribution to the general question of documentary authorship. The director’s filmography is characterized by a consistent yet self-effacing form of authorial self-inscription: Coutinho often features as an interviewer that rather than express opinions propels discourses; an interviewer that is good at listening. This mode of self-inscription characterizes him as an author who is not expressive but who is nonetheless markedly present on the screen. In Um Dia na Vida, however, Coutinho is completely absent form the image, while Últimas Conversas, on the contrary, includes a confessional prologue that moves the director from the margins to the center of his films. This article examines the ways in which these works stand out in the filmography of a director who offers new insights into the notion of cinematic authorship
Appropriate Similarity Measures for Author Cocitation Analysis
We provide a number of new insights into the methodological discussion about author cocitation analysis. We first argue that the use of the Pearson correlation for measuring the similarity between authors’ cocitation profiles is not very satisfactory. We then discuss what kind of similarity measures may be used as an alternative to the Pearson correlation. We consider three similarity measures in particular. One is the well-known cosine. The other two similarity measures have not been used before in the bibliometric literature. Finally, we show by means of an example that our findings have a high practical relevance.information science;Pearson correlation;cosine;similarity measure;author cocitation analysis
Dispelling the Myths Behind First-author Citation Counts
We conducted a full-scale evaluative citation analysis study of scholars in the XML research field to explore just how different from each other author rankings resulting from different citation counting methods actually are, and to demonstrate the capability of emerging data and tools on the Web in supporting more realistic citation counting methods. Our results contest some common arguments for the continued
use of first-author citation counts in the evaluation of scholars, such as high correlations between author rankings by first-author citation counts and other citation
counting methods, and high costs of using more realistic citation counting methods that are not well-supported by the ISI databases. It is argued that increasingly available digital full text research papers make it possible for citation analysis studies to go beyond what the ISI databases have directly supported and to employ more
sophisticated methods
koamabayili/VECTRON-author-checklist: VECTRON author checklist
We have done our best to complete the author checklist relating to the use of animals in the hut study. Note that the objective for the hut study was to evaluate the IRS treatment applications for residual efficacy against Anopheles mosquitoes, including the local An. coluzzii mosquito population. Cows were only used to attract mosquitoes into the huts and no tests were carried out directly on the cows. The author checklist is intended for use with studies where experiments are carried out on animals, which is why we have had such difficulty in completing this for the hut study, as many of the questions do not relate to how the cows were used
- …
